Related Experiment Video
Updated: Jul 10, 2026

Asthma Detection Research Based on Voice Signal Processing and Machine Learning
Published on: July 22, 2025
Improving respiratory disease detection through SSL-enhanced acoustic analysis and exercise-rest measurements
Álvaro Vera-López1, Darío Tilves-Santiago1, José Manuel Ramírez-Sánchez1
1Multimedia Technologies Group (GTM), atlanTTic Research Center, Universidade de Vigo, Vigo, Spain.
Background:
Voice analysis has emerged as a promising non-invasive approach for monitoring respiratory and systemic health conditions. However, subtle physiological alterations are often difficult to capture using recordings collected at rest. In addition, combining traditional acoustic descriptors with modern self-supervised speech representations may provide complementary information for clinical voice analysis.
Objectives:
This study evaluates a generalized screening model integrating stress-induced acoustic analysis with machine learning. We investigate how physical exertion and the fusion of traditional acoustic features with self-supervised learning embeddings (such as wav2vec 2.0 and WavLM) enhance the diagnostic sensitivity of vocal and respiratory signals. Post-Acute Sequelae of SARS-CoV-2 (PASC) is used as a case study to evaluate the proposed framework.
Methods:
Utilizing the DICOPERIA-Voice dataset (n = 154), we collected recordings of sustained vowel phonation (/a/) and voluntary coughing at two clinical moments: resting state and following a physiological stress protocol (six-minute walk and one-minute sit-to-stand tests). We employed a dual-feature extraction strategy, combining traditional acoustic biomarkers with high-dimensional Self-Supervised Learning (SSL) embeddings from wav2vec 2.0, WavLM and HuBERT. Binary classification (PASC vs. Healthy) was performed using Logistic Regression, evaluated via stratified 5-fold cross-validation.
Results:
Physical exertion significantly improved classification performance and reduced model variability across all tasks. The fusion of acoustic features, WavLM and wav2vec 2.0 achieved peak F1-scores of 82.2% for vowel phonation and 80.8% for coughing both in post-exercise conditions. A cross-task late fusion model aggregation reached the highest overall performance, with an F1-score of 87.7%.
Conclusion:
Incorporating Self-Supervised Learning representations into acoustic analysis improves the sensitivity of voice-based screening, while post-exercise measurements further enhance the robustness and consistency of classification. Together, these strategies provide a scalable and objective framework for detecting respiratory and vocal sequelae in chronic or post-viral conditions. With further validation, this approach could be integrated into routine functional assessments, offering a rapid, non-invasive adjunct to clinical decision-making.
Related Concept Videos
Respiratory System Abnormal Finding II: Palpation and Auscultation
Palpation Findings
During a respiratory assessment, palpation can reveal several vital abnormalities:
Respiratory System Abnormal Finding I: Inspection and Percussion
Inspection Findings
During an inspection, several findings may suggest the presence of respiratory distress or disease. Pursed-lip breathing, where exhalation is slowed by...
Assessment of Respiration
Subjective Assessment: Nurses interview the patient to gather information directly during the subjective assessment. It includes questions about the individual's medical history, medications, and symptoms, focusing on past respiratory conditions like asthma or COPD,...
Assessment of Airway, Skin Color, and Use of Accessory Muscles
Introduction
The initial evaluation of a patient's respiratory system...
Respiratory Assessment: Purpose and Indications
Objectives and Importance:
The primary goal of respiratory assessment is to evaluate patients at early risk of clinical deterioration. Since respiratory distress often precedes other signs of declining health, breathing patterns and sounds become a...
Physical Assessment of the Respiratory Tract IV: Auscultation
Breath Sounds
Breath sounds are categorized into vesicular, bronchovesicular, and bronchial.
