Predicting Childhood Obesity Based on Single and Multiple Well-Child Visit Data Using Machine Learning Classifiers

Pritom Kumar Mondal1, Kamrul H Foysal2, Bryan A Norman1

  • 1Department of Industrial, Manufacturing & Systems Engineering, Texas Tech University, Lubbock, TX 79409, USA.

Insights

Predicting childhood obesity early is crucial for prevention. New machine learning models accurately forecast a child's obesity risk using basic health data, aiding healthcare professionals in early intervention.

Area of Science:

  • Pediatrics
  • Public Health
  • Machine Learning

Background:

  • Childhood obesity is a significant public health issue in the U.S., leading to severe comorbidities.
  • Current methods for obesity assessment lack predictive capabilities for future risk.
  • Existing predictive models often require extensive longitudinal data and numerous variables.

Purpose of the Study:

  • To develop and evaluate novel machine learning techniques for early childhood obesity prediction.
  • To address limitations of current methods by utilizing readily available patient data.
  • To provide a decision support tool for healthcare professionals to identify at-risk children.

Main Methods:

  • Proposed three distinct machine learning models for different data availability scenarios.
  • Utilized datasets including birth BMI, gestational age, well-child visit BMI measures, and gender.
  • Applied models to predict obesity status at five years of age.

Main Results:

  • Achieved prediction accuracies of 89%, 77%, and 89% for the three distinct scenarios.
  • Demonstrated effective prediction even with limited or non-longitudinal data.
  • Models successfully categorized children into normal weight, overweight, or obese categories.

Conclusions:

  • The developed machine learning models offer a viable approach for early childhood obesity risk prediction.
  • These models can function as decision support tools, enabling timely interventions.
  • Early prediction and intervention can mitigate long-term health complications associated with childhood obesity.

Related Concept Videos

Obesity01:24

Obesity

The Body Mass Index (BMI) is a numerical value derived from a person's weight and height, used to categorize individuals into weight ranges. It is calculated using the formula: weight in kilograms divided by height in meters squared. Obesity is a health condition characterized by excessive accumulation of adipose tissue that poses health risks, often diagnosed with a BMI ≥ 30. This excess fat storage occurs when surplus dietary calories are converted into triglycerides and stored in...
569
Classification of Illness01:17

Classification of Illness

The meaning of illness is individualized to each person who experiences an alteration in health. In contrast, disease is a medical term indicating a pathological change in the structure and function of the body or mind. It is a condition that has specific symptoms and boundaries.
An illness is a response to a disease in which the person's level of functioning is changed compared with a previous level. The general classification of illness includes acute and chronic.
Acute illness is severe...
7.7K
Steps in Outbreak Investigation01:18

Steps in Outbreak Investigation

In the ever-evolving field of public health, statistical analysis serves as a cornerstone for understanding and managing disease outbreaks. By leveraging various statistical tools, health professionals can predict potential outbreaks, analyze ongoing situations, and devise effective responses to mitigate impact. For that to happen, there are a few possible stages of the analysis:
166
Residuals and Least-Squares Property01:11

Residuals and Least-Squares Property

The vertical distance between the actual value of y and the estimated value of y. In other words, it measures the vertical distance between the actual data point and the predicted point on the line
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
7.8K