Related Experiment Video
Updated: Sep 10, 2025

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
Published on: September 27, 2024
Language-specific phonetic realisation of stop voicing contrasts in English and Japanese synthesised speech
James Tanner1, Yasuaki Shinohara2, Faith Chiu1
1English Language and Linguistics, University of Glasgow, Glasgow, G12 8QQ, United Kingdom.
Abstract:
Speech synthesis has improved dramatically over recent years, enabled by large datasets and advances in neural network architectures. Little is known, however, about how synthesised speech patterns are realized from a phonetic perspective. By synthesising speech in two languages with differing implementations of stop voicing, we observe that synthesised speech broadly follows expected patterns for each language, though partially diverges for specific segments. Synthesising speakers into the opposing language also results in stops similar to target language distributions. These findings demonstrate the capability of speech synthesis models to encode phonetic information and further motivate questions regarding the phonetics of synthesised speech.
More Related Videos
Related Concept Videos
Components of Language
Air-entraining Agents
Auditory Perception
Perceiving Loudness, Pitch, and Location
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
Double Resonance Techniques: Overview
Spin decoupling is usually achieved by...
Tip-of-the-Tongue Phenomenon

