Hearing a face: cross-modal speaker matching using isolated visible speech
Lawrence D Rosenblum1, Nicolas M Smith, Sarah M Nichols
1Department of Psychology, University of California, Riverside, CA 92521, USA. rosenblu@citrus.ucr.edu
Abstract:
An experiment was performed to test whether cross-modal speaker matches could be made using isolated visible speech movement information. Visible speech movements were isolated using a point-light technique. In five conditions, subjects were asked to match a voice to one of two (unimodal) speaking point-light faces on the basis of speaker identity. Two of these conditions were designed to maintain the idiosyncratic speech dynamics of the speakers, whereas three of the conditions deleted or distorted the dynamics in various ways. Some of these conditions also equated video frames across dynamically correct and distorted movements. The results revealed generally better matching performance in the conditions that maintained the correct speech dynamics than in those conditions that did not, despite containing exactly the same video frames. The results suggest that visible speech movements themselves can support cross-modal speaker matching.
Related Concept Videos
Perceiving Loudness, Pitch, and Location
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by identifying...
Facial Feedback Hypothesis
Auditory Perception
Hearing
Prosopagnosia
Nonconscious Mimicry

