Related Experiment Video
Updated: Apr 26, 2026

09:09
Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
Published on: September 27, 2024
1.1K
Speaker Identification Using Voice Quality Features: A Psychoacoustic and Machine Learning Approach
Homa Asadi1, Batool Alinezhad1, Razieh Zare1
1Department of Linguistics, Faculty of Foreign Languages, University of Isfahan, Isfahan, Iran.
Journal of Voice : Official Journal of the Voice Foundation
|April 24, 2026
Summary
Psychoacoustic voice features effectively identify Persian speakers. Dynamic features aid male speaker recognition, while stable features are key for females, showing sex-dependent vocal identity cues.
Area of Science:
- Phonetics and Speech Science
- Acoustic Analysis
- Biometrics
Background:
- Speaker identification relies on unique voice characteristics.
- Psychoacoustic models offer a framework for analyzing voice quality.
- Understanding sex-specific voice features is crucial for accurate speaker recognition.
Purpose of the Study:
- To assess the effectiveness of psychoacoustic voice features for closed-set speaker identification.
- To determine the discriminative power of static and dynamic voice features.
- To identify key features for distinguishing individual Persian speakers.
Main Methods:
- Collected speech recordings from 120 native Persian speakers (60 male, 60 female).
- Extracted 26 psychoacoustic voice features (13 static, 13 dynamic using coefficient of variation).
- Utilized a Random Forest classifier for speaker separability analysis.
Main Results:
- Combined features achieved 90.15% accuracy for males and 90.74% for females.
- Static features were more effective for females (91.24%) than dynamic features (76.52%).
- Dynamic features improved accuracy for males (90.15%) compared to static alone (89.53%).
- Key features differed by sex: males (H1*-H2*, CoVH1*-H2*, f0), females (f0, CoVf0, F3).
Conclusions:
- Psychoacoustic voice features are effective for Persian speaker identification.
- The utility of dynamic voice features for speaker discrimination is sex-dependent.
- Temporal variability enhances male speaker recognition, while stable features are more critical for females.
- Findings inform forensic phonetics and speaker recognition system development.
Related Concept Videos
Perceiving Loudness, Pitch, and Location
1.3K
The human brain perceives pitch through two primary mechanisms reflected in place theory and frequency theory. Each mechanism describes how sound waves are interpreted as specific pitches by the brain, offering insights into the intricate processes of auditory perception.
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
1.3K
Classification of Signals
1.5K
In signal processing, signals are classified based on various characteristics: continuous-time versus discrete-time, periodic versus aperiodic, analog versus digital, and causal versus noncausal. Each category highlights distinct properties crucial for understanding and manipulating signals.
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
1.5K
Auditory Perception
1.5K
The auditory system is essential for sound perception, utilizing various critical structures. When sound waves enter the outer ear, they travel through the ear canal and cause the eardrum to vibrate. These vibrations are then transmitted to the middle ear, where three tiny bones – the malleus, incus, and stapes – amplify the sound. This amplification is crucial, as it ensures that the sound vibrations are strong enough to be conveyed to the inner ear. These vibrations then reach the...
1.5K

