Validation of an embedding-based forensic voice comparison system using short speech samples in Chinese languages

Bruce Xiao Wang1, Cuiling Zhang2, Ricky K W Chan3

  • 1Department of English and Communication, The Hong Kong Polytechnic University, Hong Kong, China.

Summary

Forensic voice comparison performance using ECAPA-TDNN x-vectors significantly improves with speech duration up to 10 seconds. Calibration size had a lesser impact, with Hong Kong Cantonese outperforming Northeast Mandarin.

Related Concept Videos