Related Experiment Video
Updated: Aug 23, 2026

Testing Sensory and Multisensory Function in Children with Autism Spectrum Disorder
Published on: April 22, 2015
[Intermodal timing cues for audio-visual speech recognition]
Masahiro Hashimoto1, Masaharu Kumashiro
1Bio-information Research Center, University of Occupational and Environmental Health, Yahatanishi-ku, Kitakyushu 807-8555, Japan.
Abstract:
The purpose of this study was to investigate the limitations of lip-reading advantages for Japanese young adults by desynchronizing visual and auditory information in speech. In the experiment, audio-visual speech stimuli were presented under the six test conditions: audio-alone, and audio-visually with either 0, 60, 120, 240 or 480 ms of audio delay. The stimuli were the video recordings of a face of a female Japanese speaking long and short Japanese sentences. The intelligibility of the audio-visual stimuli was measured as a function of audio delays in sixteen untrained young subjects. Speech intelligibility under the audio-delay condition of less than 120 ms was significantly better than that under the audio-alone condition. On the other hand, the delay of 120 ms corresponded to the mean mora duration measured for the audio stimuli. The results implied that audio delays of up to 120 ms would not disrupt lip-reading advantage, because visual and auditory information in speech seemed to be integrated on a syllabic time scale. Potential applications of this research include noisy workplace in which a worker must extract relevant speech from all the other competing noises.
Related Concept Videos
Non-Verbal Cues
Chunking and Rehearsal in Sensory Memory
Auditory Pathway
When viewed cross-sectionally, the cochlea reveals the scala vestibuli and scala tympani flanking the...