Quantifying and Improving the Performance of Speech Recognition Systems on Dysphonic Speech.

Julio C Hidalgo Lopez1, Shelly Sandeep1, MaKayla Wright2

  • 1Emory University School of Medicine, Atlanta, Georgia, USA.

Summary

Current speech recognition systems struggle with dysphonic speech. A custom model significantly improved accuracy for conditions like spasmodic dysphonia, achieving over 96% performance on dysphonic voices.

Related Concept Videos

Frequency-Domain Interpretation of PD Control01:24

Frequency-Domain Interpretation of PD Control

Proportional-Derivative (PD) controllers are widely used in fan control systems to improve stability and performance. A fan control system can be effectively represented using a Bode plot to illustrate the impact of a PD controller through its transfer function. The Bode plot visually conveys how PD control modifies the fan's response across various frequencies, providing a frequency domain interpretation of the controller's behavior.
The proportional control gain, combined with the...
155
Equipments Used To Measure Blood Pressure01:30

Equipments Used To Measure Blood Pressure

Direct Method
This invasive approach involves cannulating a peripheral artery. During each cardiac contraction, pressure generates mechanical motion within the catheter, transmitted through rigid, fluid-filled tubing to a transducer. This transducer converts mechanical motion into electrical signals displayed as waveforms on a monitor. An automatic flushing system prevents blood backflow. Due to the potential risk of unexpected arterial blood loss, this method is primarily used in intensive...
1.1K
Improving Translational Accuracy02:07

Improving Translational Accuracy

Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
11.7K