Related Experiment Video
Updated: Jan 9, 2026

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
Published on: September 27, 2024
Between- and within-speaker variability of voiceless fricatives in Persian
Homa Asadi1, Batool Alinezhad1, Volker Dellwo2
1Department of Linguistics, University of Isfahan, Azadi Square, University Street, Isfahan 81746-73441, Iran.
Abstract:
Fricatives vary acoustically across languages and individuals, with speaker variability shaped by both phonetic and non-phonetic factors. This study examined between- and within-speaker variability in Persian voiceless fricatives (/f/, /s/, /ʃ/, /x/) and how linguistic environments, such as syllable position and lexical stress, affect this variability. A gender-balanced sample of 24 Persian speakers was recorded in two sessions, 1-2 two weeks apart. Acoustic analysis targeted the first four spectral moments and duration. Results showed that center of gravity captured the greatest between-speaker variability, followed by standard deviation, skewness, duration, and kurtosis. Across segments, the alveolar /s/ exhibited the highest speaker-specificity, followed by /ʃ/, /f/, and /x/. Gender-based patterns emerged: for males, the center of gravity and skewness of /s/ were most discriminative, whereas for females, the center of gravity and standard deviation of /ʃ/ were most effective. The labiodental /f/ showed some speaker-specific characteristics only in the male group. Voiceless fricatives in syllable-initial positions demonstrated more speaker-specificity, while lexical stress did not impact between-speaker variability. Results also highlight cross-linguistic differences in the acoustic cues most effective for speaker differentiation and demonstrate that optimal features can vary across speaker populations. Adaptive algorithms are therefore crucial for improving forensic speaker comparison.
Related Concept Videos
Variability: Analysis
The range is a simple measure of variability, indicating the difference between the highest and...
IR Spectrum Peak Splitting: Symmetric vs Asymmetric Vibrations
Testing a Claim about Mean: Unknown Population SD
Estimating a population mean requires the samples to be approximately normally distributed. The data should be collected from the randomly selected samples having no sampling bias. There is no specific requirement for sample size. But if the sample size is less than 30, and we don't know the population standard deviation, a different approach is used;...
Larynx
Anatomy of the Larynx
The larynx consists of various components, including cartilage, muscles, and vocal cords. Its structure includes three large unpaired cartilages—the thyroid, cricoid, and epiglottis—and three smaller paired cartilages—the arytenoids,...
What is Variation?
The range, standard deviation, standard error, and variance are the different measures of variation.
Range: The range is the difference between its maximum and...
Interference: Path Lengths
Two special sources may be considered when they are in phase. This can be easily achieved by feeding the two sources from the same source. An example would be synchronizing the two speakers by feeding them with the same source, such as the sound waves produced by a tuning fork. This setup ensures that the two sources have the same frequency and are...

