Related Experiment Video
Updated: Sep 30, 2026

Minimally Invasive Murine Laryngoscopy for Close-Up Imaging of Laryngeal Motion During Breathing and Swallowing
Published on: December 1, 2023
Machine Learning for Non-invasive Differentiation of Glottic Carcinoma Versus Laryngeal Immobility
Clémence Forges1,2,3, Robin Baudouin1,3,4,5, Stanislas Nicolleau1,4,6
1Department of Otolaryngology-Head & Neck Surgery, Foch Hospital, Suresnes, France.
Objective:
Voice analysis for medical diagnostics is advancing rapidly with artificial intelligence and machine learning. Non-invasive voice analysis could optimize diagnostic procedures in ear, nose, and throat (ENT), supplementing endoscopic evaluations and guiding treatment strategies. This study aims to distinguish early glottic carcinomas (EGCs) from unilateral laryngeal immobility (ULI) in dysphonic patients using machine learning.
Study Design:
A machine learning model was developed in Python, incorporating data segmentation, normalization, hyperparameter tuning, and data augmentation via parameter transformation and using Synthetic Minority Over-sampling Technique (SMOTE) to balance class distributions. A metamodel combining multiple algorithms (Random Forest Classifier, Support Vector Machines, and Extreme Gradient Boosting) enhanced predictive accuracy.
Setting:
Tertiary Medical Center (laryngology unit, ENT Department).
Methods:
Population is dysphonic patients diagnosed with ULI or EGC, and asymptomatic controls. Patients with other voice-affecting conditions were excluded. Sixty-three pathologic cases (ULI: 38, EGC: 25) and 40 controls. Data augmentation and rebalance expanded the data set to 650 samples. Voice recordings of sustained vowels were analyzed for vocal parameters (fundamental frequency, jitter, and shimmer) using machine learning. High-quality audio equipment was used. Model performance was evaluated by precision, recall, and area under receiver operating characteristic curve (AUC) in differentiating EGC from ULI.
Results:
The model achieved 72% accuracy in distinguishing ULI from EGC, validated through tests.
Conclusion:
Machine learning shows strong potential for distinguishing EGC from ULI. This non-invasive method could improve diagnostic accuracy and enable earlier intervention, especially in regions with limited specialist access. Further refinement could enhance its role in public health.
Related Concept Videos
Larynx
Anatomy of the Larynx
The larynx consists of various components, including cartilage, muscles, and vocal cords. Its structure includes three large unpaired cartilages—the thyroid, cricoid, and epiglottis—and three smaller paired cartilages—the arytenoids, corniculates, and...
Cardiopulmonary Resuscitation V: Advanced Airway Management Techniques

