Related Experiment Video
Updated: Jul 12, 2026

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
Single-label and multi-label classification for disease recognition with special consideration of comorbidities
Sophie Schmiegel1, Hannah Marchi1, Marvin-Hendrik Röchter2,3
1Faculty of Business Administration and Economics, Data Science Group, Bielefeld University, Universitätsstraße 25, Bielefeld, 33615, Germany.
Background:
Certain diseases require rapid treatment to avoid long-term consequences for patients. However, they may be difficult to recognize, especially if the symptoms are ambiguous and compatible with multiple possible diagnoses. Completing all necessary examinations often takes time, thereby prolonging patient suffering. Data-driven approaches, such as single-label classification (SLC) and multi-label classification (MLC), can help accelerate the diagnostic process and improve accuracy.
Methods:
Comparing SLC and MLC allows us to investigate whether disease recognition benefits from considering comorbidities. In this context, we aim to provide a conceptual framing of (differences between) the two approaches in model formulation, decision spaces and handling of class imbalance. To empirically assess their performance, we conduct a case study applying SLC and MLC to data from chronic pain patients.
Results:
Our analysis yields an ambiguous picture of whether incorporating comorbidities improves disease recognition. The suitability of SLC and MLC is determined by multiple factors, notably the dependency structure among diseases and between diseases and covariates, as well as by data characteristics such as class imbalance. Consequently, it is of high importance to address each individual problem in an individual manner, especially when the implications can be significant, such as in patient referrals.
Conclusion:
Our study highlights the importance of considering the specific characteristics of the data when selecting an appropriate classification approach for disease recognition and beyond. We illustrate how underlying assumptions, learning structures, and strategies for handling class imbalance shape predictive performance. Our results thus lead to the point that computational prediction tools should serve as decision support for treating physicians, but never replace their clinical judgment.
Related Concept Videos
Classification of Illness
An illness is a response to a disease in which the person's level of functioning is changed compared with a previous level. The general classification of illness includes acute and chronic.
Acute illness is severe and...
Heart Failure IV: Classification and Diagnostic Evaluation
Classification of Systems-II
Classification of Systems-I
Homogeneity dictates that if an input x(t) is multiplied by a constant c, the output y(t) is multiplied by the same constant. Mathematically, this is expressed as:
Cardiomyopathy I: Introduction and Classification
Genomics
