Related Experiment Video
Updated: Jan 18, 2026

Dynamic Digital Biomarkers of Motor and Cognitive Function in Parkinson's Disease
Published on: July 24, 2019
Harnessing Clinical and Biochemical Data for Personalized Cardiovascular Risk Prediction: a Machine Learning Approach
Joyeta Ghosh1, Tinni Chaudhuri2, Jose Arturo Molina Mora3
1Department of Dietetics and Applied Nutrition, Amity Institute of Applied Sciences (AIAS), Amity University - Kolkata Campus, Kolkata, West Bengal, India.
Background:
Cardiovascular disease (CVD) is a leading cause of morbidity and mortality among postmenopausal women in rural India, where healthcare resources remain limited.
Objectives:
This study aimed to leverage artificial intelligence (AI) and machine learning (ML) approaches to predict CVD risk in rural elderly women, identify key clinical predictors, and assess model performance using interpretable AI tools.
Methods:
This observational cross-sectional study was conducted in Singur Block (West Bengal) and Amdanga Block (North 24 Parganas District) between March 2014 and August 2018. Data from 458 rural postmenopausal women were analyzed. The outcome variable was the presence or absence of elevated cardiovascular disease risk, defined using composite International Diabetes Federation and American Heart Association criteria. Predictors included waist circumference, blood pressure, fasting blood glucose, HDL cholesterol, triglycerides, and vitamin D concentrations. Seven ML models [Random Forest, Gradient Boosting, Ensemble (Voting Classifier), Extra Trees, Support Vector Machine, Neural Network, and Logistic Regression] were developed and compared. Model evaluation employed 5-fold cross-validation with metrics including accuracy, AUC, precision, recall, and F1 score.
Results:
Among the 458 participants, 171 (37.3%) exhibited elevated CVD risk. The Random Forest model achieved an accuracy of 98.91% (95% CI: 97.8%, 99.6%), whereas eXtreme Gradient Boosting (XGBoost) demonstrated comparable performance with an AUC of 0.998 (95% CI: 0.993, 1.000), precision of 97.2%, and recall of 98.3%. Feature-importance analysis revealed waist circumference, blood pressure, and fasting glucose as the strongest predictors, with HDL cholesterol and vitamin D contributing modestly but significantly.
Conclusions:
ML models-particularly Random Forest and XGBoost-demonstrated high accuracy and interpretability in predicting CVD risk among rural postmenopausal women. These findings highlight the potential of AI-driven, low-cost predictive tools for early CVD risk detection and personalized preventive healthcare in resource-limited rural settings.
More Related Videos
07:51Hydra, a Computer-Based Platform for Aiding Clinicians in Cardiovascular Analysis and Diagnosis
Published on: September 26, 2018
04:09Predicting Treatment Response to Image-Guided Therapies Using Machine Learning: An Example for Trans-Arterial Treatment of Hepatocellular Carcinoma
Published on: October 10, 2018
Related Concept Videos
Blood Studies for Cardiovascular System I: Cardiac Biomarkers
The essential diagnostic tools for detecting myocardial necrosis and monitoring individuals suspected of having acute coronary syndrome (ACS) include:
Troponins
Troponins, particularly cardiac troponins I and T, are the most precise and sensitive markers of myocardial injury. They are detectable within 4-6 hours of myocardial injury and remain...
Combination Therapies and Personalized Medicine
The combination of the drug acetazolamide and sulforaphane is a good example of combination therapy to treat cancer. The cells in the interior of a large tumor often die due to the hypoxic and...
Blood Studies for Cardiovascular System II: CRP, Hcy, and Cardiac Natriuretic Peptide Markers
These markers indicate stress or strain on the heart muscle:
Natriuretic Peptides (BNP)
Cardiac myocytes produce these hormones in response to ventricular stretching...
Coronary Artery Disease I: Introduction
Atherosclerosis III: Management
Coronary Artery Disease IV: Preventive Measures