Related Experiment Video
Updated: Sep 11, 2025

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
Published on: September 27, 2024
Improving Voice Spoofing Detection Through Extensive Analysis of Multicepstral Feature Reduction
Leonardo Mendes de Souza1, Rodrigo Capobianco Guido1, Rodrigo Colnago Contreras2
1Department of Computer Science and Statistics, Institute of Biosciences, Letters and Exact Sciences, São Paulo State University, São José do Rio Preto 15054-000, SP, Brazil.
This study enhances voice biometric security by using dimensionality reduction techniques to detect AI-generated spoofing attacks. Applying methods like PCA and Random Forest feature importance significantly improves spoof detection accuracy, achieving an Equal Error Rate (EER) of around 10%.
Area of Science:
- Biometrics and Security
- Artificial Intelligence and Machine Learning
- Signal Processing
Background:
- Voice biometric systems are crucial for security but vulnerable to AI-powered spoofing attacks.
- Realistic synthetic speech poses a significant threat to current voice authentication methods.
- Developing robust defenses against sophisticated spoofing is essential for maintaining system integrity.
Purpose of the Study:
- To explore and evaluate various dimensionality reduction strategies for identifying spoofed voice signals.
- To enhance the effectiveness of supervised machine learning models in voice anti-spoofing.
- To assess the performance gains achieved through dimensionality reduction in voice biometric security.
Main Methods:
- Extraction of multi-cepstral features from voice signals.
- Application of diverse dimensionality reduction techniques: PCA, SVD, ANOVA F-value, Mutual Information, RFE, LASSO, Random Forest importance, Permutation Importance.
- Empirical evaluation using the ASVSpoof 2017 v2.0 dataset and Equal Error Rate (EER) metric.
Main Results:
- Dimensionality reduction methods significantly improved the performance of spoof detection.
- Achieved an Equal Error Rate (EER) of approximately 10% on the ASVSpoof 2017 v2.0 dataset.
- Demonstrated the effectiveness of feature selection and reduction in combating voice spoofing.
Conclusions:
- Dimensionality reduction is a valuable approach for enhancing voice biometric security against AI-driven spoofing.
- The proposed framework offers improved accuracy and robustness for voice authentication systems.
- Further research into advanced feature engineering and reduction techniques can bolster defenses against evolving threats.
Related Concept Videos
Downsampling
The Fourier transform of the decimated sequence reveals a combination of scaled and shifted versions of the original spectrum. This...
IR Spectrum Peak Splitting: Symmetric vs Asymmetric Vibrations
Improving Translational Accuracy
Linear Approximation in Frequency Domain
In contrast, nonlinear systems do not inherently possess these properties. However, for small deviations around an operating point, a nonlinear system can often be approximated as linear....
Masking and Demasking Agents
There are many masking agents, such as cyanide, fluoride, triethanolamine, thiourea, and 2,3-bis(sulfanyl)propan-1-ol (formerly 2,3-dimercapto-1-propanol), with the masking agent chosen based on...
Double Resonance Techniques: Overview
Spin decoupling is usually achieved by...

