Related Experiment Video
Updated: Mar 24, 2026

09:09
Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
Published on: September 27, 2024
965
Differences in coda voicing trigger changes in gestural timing: A test case from the American English diphthong /aɪ/
1Department of Linguistics, University of Wisconsin, Milwaukee P.O. Box 413, Milwaukee, Wisconsin 53211-0413, U.S.A.
Summary
The timing of articulatory gestures for speech sounds changes based on whether the following consonant is voiced or voiceless. This timing difference impacts how we perceive speech sounds, revealing sensitivity to underlying articulatory timing.
Area of Science:
- Phonetics and Phonology
- Speech Production and Perception
- Articulatory Linguistics
Background:
- Vowel duration and spectral characteristics differ before voiceless versus voiced consonants.
- Existing theories propose various explanations for these acoustic differences.
- This study explores a unified source: the temporal reorganization of articulatory gestures.
Purpose of the Study:
- To test the hypothesis that acoustic differences in vowels before codas stem from temporal reorganization of articulatory gestures.
- To investigate the role of temporal gestural distance in producing and perceiving voicing contrasts.
- To examine the American English diphthong /aɪ/ as a case study.
Main Methods:
- Experiment 1: Measured the ratio of nucleus-to-offglide duration for the diphthong /aɪ/ in production before voiceless and voiced codas.
- Experiment 1: Analyzed the effect of speech rate and phrasal position on this ratio.
- Experiment 2: Assessed listeners' perception of diphthongs with contextually incongruent nucleus-to-offglide ratios before voiceless codas.
Main Results:
- Production: The nucleus-to-offglide duration ratio was smaller before voiceless codas than before voiced codas, consistent across speakers and conditions.
- Perception: Listeners showed delayed word identification when diphthongs had contextually incongruent ratios, even with intact voicing cues.
- These findings indicate sensitivity to the temporal gestural origins of voicing differences.
Conclusions:
- The voicing contrast in codas influences the temporal organization of preceding articulatory gestures.
- Gestures are temporally closer before voiceless codas and further apart before voiced codas.
- Listeners are sensitive to these timing variations, supporting a gestural account of phonetic variation.
More Related Videos
Related Concept Videos
The Auditory Ossicles
3.7K
The auditory ossicles of the middle ear transmit sounds from the air as vibrations to the fluid-filled cochlea. The auditory ossicles consist of two malleus (hammer) bones, two incus (anvil) bones, and two stapes (stirrups), one on each side. These bones develop during the fetal stage and are the ones to ossify first. They are fully mature at birth and do not grow afterward.
The aptly named stapes look very much like a stirrup. The three ossicles are unique to mammals, and each plays a role in...
The aptly named stapes look very much like a stirrup. The three ossicles are unique to mammals, and each plays a role in...
3.7K
Doppler Effect - II
5.1K
The Doppler effect has several practical, real-world applications. For instance, meteorologists use Doppler radars to interpret weather events based on the Doppler effect. Typically, a transmitter emits radio waves at a specific frequency toward the sky from a weather station. The radio waves bounce off the clouds and precipitation and travel back to the weather station. The radio frequency of the waves reflected back to the station appears to decrease if the clouds or precipitation are moving...
5.1K
Deglutition
7.5K
Swallowing, otherwise known as deglutition, facilitates the transport of food from the mouth to the stomach. It is a multifaceted process that involves both the tongue and the muscles of the throat and esophagus. Saliva and mucus aid in this process, which takes approximately 4 to 8 seconds for semi-solid or solid food and around 1 second for liquids or very soft food.
Swallowing can be divided into three stages: the voluntary phase, the pharyngeal phase, and the esophageal phase. Although the...
Swallowing can be divided into three stages: the voluntary phase, the pharyngeal phase, and the esophageal phase. Although the...
7.5K
Larynx
6.0K
The human larynx, often referred to as the voice box, is an intricate organ located in the neck. It serves as a pathway for air to enter the lungs during respiration and is an essential component of voice production.
Anatomy of the Larynx
The larynx consists of various components, including cartilage, muscles, and vocal cords. Its structure includes three large unpaired cartilages—the thyroid, cricoid, and epiglottis—and three smaller paired cartilages—the arytenoids,...
Anatomy of the Larynx
The larynx consists of various components, including cartilage, muscles, and vocal cords. Its structure includes three large unpaired cartilages—the thyroid, cricoid, and epiglottis—and three smaller paired cartilages—the arytenoids,...
6.0K
Tip-of-the-Tongue Phenomenon
628
The tip-of-the-tongue (TOT) phenomenon is a cognitive experience characterized by a temporary inability to retrieve specific information from memory despite having a strong feeling of knowing the information. Although individuals cannot access the target word or detail, they frequently recall related elements, such as its initial letter, syllable count, or context. This partial retrieval often causes frustration, as one might recognize a familiar face or know that a name starts with a specific...
628
Hearing
58.6K
When we hear a sound, our nervous system is detecting sound waves—pressure waves of mechanical energy traveling through a medium. The frequency of the wave is perceived as pitch, while the amplitude is perceived as loudness.
58.6K

