Related Experiment Video
Updated: Aug 8, 2025

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
HiPerMAb: a tool for judging the potential of small sample size biomarker pilot studies
Amani Al-Mekhlafi1,2, Frank Klawonn3
1Department of Biostatistics, Helmholtz Centre for Infection Research, Braunschweig, Germany.
Abstract:
Common statistical approaches are not designed to deal with so-called "short fat data" in biomarker pilot studies, where the number of biomarker candidates exceeds the sample size by magnitudes. High-throughput technologies for omics data enable the measurement of ten thousands and more biomarker candidates for specific diseases or states of a disease. Due to the limited availability of study participants, ethical reasons and high costs for sample processing and analysis researchers often prefer to start with a small sample size pilot study in order to judge the potential of finding biomarkers that enable - usually in combination - a sufficiently reliable classification of the disease state under consideration. We developed a user-friendly tool, called HiPerMAb that allows to evaluate pilot studies based on performance measures like multiclass AUC, entropy, area above the cost curve, hypervolume under manifold, and misclassification rate using Monte-Carlo simulations to compute the p-values and confidence intervals. The number of "good" biomarker candidates is compared to the expected number of "good" biomarker candidates in a data set with no association to the considered disease states. This allows judging the potential in the pilot study even if statistical tests with correction for multiple testing fail to provide any hint of significance.
More Related Videos
08:14MicroRNA Based Liquid Biopsy: The Experience of the Plasma miRNA Signature Classifier MSC for Lung Cancer Screening
Published on: October 26, 2017
07:41Performing Data Mining And Integrative Analysis Of Biomarker in Breast Cancer Using Multiple Publicly Accessible Databases
Published on: May 17, 2019
Related Concept Videos
McNemar's Test
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...