Related Experiment Video
Updated: May 6, 2026

Automated, Long-term Behavioral Assay for Cognitive Functions in Multiple Genetic Models of Alzheimer's Disease, Using IntelliCage
Published on: August 4, 2018
Interpretable machine learning for cognitive impairment screening: Development and external validation of a clinical
Kang Chen1, Guran Yu1, Hao Li2
1Department of Neurology, Jiangsu Province Hospital of Chinese Medicine, the Affiliated Hospital of Nanjing University of Chinese Medicine, Nanjing, Jiangsu Province, China.
Background:
Cognitive impairment in older adults poses a growing public health challenge, yet accessible screening tools remain limited. We aimed to develop and validate an interpretable machine learning model for cognitive impairment prediction by routinely collecting clinical data.
Methods:
We analyzed 1061 participants from the U.S. National Health and Nutrition Examination Survey (NHANES 2011-2014). Feature selection combined multivariable regression, restricted cubic splines, and the Boruta algorithm to identify 40 clinical, demographic, and socioeconomic variables. Twelve machine learning models (including Support Vector Machine (SVM), Extreme Gradient Boosting (XGBoost), and Random Forest (RF)) were trained and externally validated on NHANES 2001-2002 (n = 531). Model performance was evaluated by area under the receiver operating characteristic curve (AUC-ROC), calibration (Brier score), accuracy, and sensitivity. Additionally, an assessment of fairness was conducted across racial subgroups. Interpretability was enhanced via SHapley Additive exPlanations (SHAP).
Results:
The SVM model demonstrated optimal generalizability, achieving an external validation AUC of 0.8265 (95 %CI: 0.7867-0.8582) with sustained calibration (Brier score = 0.1703). Subgroup analyses showed no statistically significant AUC differences (all P > 0.05). SHAP analysis identified socioeconomic factors, systemic inflammation indices, and metabolic markers as key predictors.
Limitations:
Generalizability may be limited to U.S. populations, and unmeasured biomarkers (e.g., amyloid-β) could affect prediction accuracy. Subgroup analyses for minorities were constrained by sample size.
Conclusion:
Our interpretable prediction strategy enables rapid cognitive risk assessment using routine clinical data, providing a cost-effective decision support tool adaptable to electronic health record systems.
More Related Videos
Related Concept Videos
Data Validation
Nursing assessment guides are generally based on holistic models rather than medical...
Analysis Methods of Pharmacokinetic Data: Model and Model-Independent Approaches
The model approach uses mathematical models to describe changes in drug concentration over time. Pharmacokinetic models help characterize drug behavior in patients, predict drug concentration in the body fluids, calculate optimum dosage regimens, and evaluate the risk of toxicity. However, ensuring that the model fits the experimental data accurately...
Model Approaches for Pharmacokinetic Data: Distributed Parameter Models
The distributed parameter models are specifically designed to account for variations and differences in some drug classes. This model is particularly useful for assessing regional concentrations of anticancer or...
Pharmacokinetic Models: Comparison and Selection Criterion
Physiological models take a detailed approach by considering specific molecular processes. They can predict drug distribution, metabolism, and elimination changes, providing a comprehensive understanding of how drugs interact with the body.
Statistical Methods for Analyzing Epidemiological Data

