Related Experiment Video
Updated: Jan 28, 2026

A Machine Learning Approach to Design an Efficient Selective Screening of Mild Cognitive Impairment
Published on: January 11, 2020
Resolving Interpretation Challenges in Machine Learning Feature Selection With an Iterative Approach in Biomedical
Jörn Lötsch1,2,3, André Himmelspach1, Dario Kringel1
1Faculty of Medicine, Goethe University, Institute of Clinical Pharmacology, Frankfurt am Main, Germany.
Background:
Machine learning (ML) is increasingly used to analyse pain-related data, emphasising how well variables classify individuals, that is, training an algorithm to assign people to predefined groups such as high versus low pain sensitivity, rather than focusing on p-values. A challenge arises when accurate classification persists after removing variables identified as important by feature-selection methods. This creates uncertainty about which factors are genuinely relevant to the trait of interest, as classification information may still reside in the remaining features.
Methods:
An iterative ML framework is presented that repeatedly tests groups of variables, combining two established feature-selection techniques with several classification algorithms. The approach was applied to three datasets, two assessing pain traits and one artificial, and compared with classical statistical methods, including logistic regression.
Results:
The iterative process clarified which variables were truly relevant for classification by assessing whether unselected features could still discriminate individuals. When they could not, selected variables became more interpretable in a biological context. Combining multiple ML approaches improved feature selection, addressed multicollinearity and enhanced robustness across models. Logistic regression sometimes required preselected inputs or missed known relevant variables. Variation in model performance increased interpretive complexity.
Conclusions:
ML-based feature selection broadens methodological options for identifying trait-relevant variables. Iterating through variable sets supports transparent, replicable inference. ML can help identify variables related to pain traits, but selected features should not be assumed uniquely important. Testing unselected variables remains essential, as their failure to predict outcomes may reflect algorithmic limitations rather than definitive trait exclusivity.
Significance Statement:
This study presents an iterative machine learning framework that improves the identification of trait-relevant features in biomedical pain data. This framework reduces ambiguity in feature selection and clarifies interpretation, helping to distinguish robust, meaningful predictors from coincidental ones. This approach enhances the interpretation and transparency of machine learning analyses in pain research and related biomedical fields.
More Related Videos
09:34A Virtual Machine Platform for Non-Computer Professionals for Using Deep Learning to Classify Biological Sequences of Metagenomic Data
Published on: September 25, 2021
07:15Machine Learning Algorithms for Early Detection of Bone Metastases in an Experimental Rat Model
Published on: August 16, 2020
Related Concept Videos
Bioequivalence Data: Statistical Interpretation
Selected Data About Geographic Locations
Model Approaches for Pharmacokinetic Data: Compartment Models
Two primary types of compartment models are recognized: mammillary and catenary. The more...
Model Approaches for Pharmacokinetic Data: Physiological Models
Interpreting R Charts
An R chart plots the range of subsets of measurements collected from a process. Each point on the chart represents the range—defined as the difference between the maximum and minimum...
Machines
A free-body diagram of the...