A semi-supervised algorithm for improving the consistency of crowdsourced datasets: The COVID-19 case study on

Lara Orlandic1, Tomas Teijeiro2, David Atienza1

  • 1Embedded Systems Laboratory (ESL), EPFL, Lausanne, Switzerland.

Summary

Semi-supervised learning (SSL) enhances cough audio classification by improving data consistency for COVID-19 detection and cough characterization. This method aggregates expert knowledge, creating a more reliable dataset for training diagnostic models.

Related Concept Videos

Classification of Illness01:17

Classification of Illness

The meaning of illness is individualized to each person who experiences an alteration in health. In contrast, disease is a medical term indicating a pathological change in the structure and function of the body or mind. It is a condition that has specific symptoms and boundaries.
An illness is a response to a disease in which the person's level of functioning is changed compared with a previous level. The general classification of illness includes acute and chronic.
Acute illness is severe...
7.6K
Common Respiratory Disorders01:31

Common Respiratory Disorders

Respiratory disorders, a prevalent health concern globally, are generally divided into two primary categories: upper and lower respiratory tract disorders. The categorization is based on the area of the respiratory system they affect.
Upper respiratory disorders impact the airways above the vocal cords, encompassing areas like the nose, sinuses, and throat. Various conditions fall under this category, including the common cold and allergic rhinitis. These disorders can stem from several causes,...
667
Residuals and Least-Squares Property01:11

Residuals and Least-Squares Property

The vertical distance between the actual value of y and the estimated value of y. In other words, it measures the vertical distance between the actual data point and the predicted point on the line
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
7.4K