Evaluating Ensemble-Based Machine Learning Models for Diagnosing Pediatric Acute Appendicitis: Insights from a

Zeynep Kucukakcali1, Sami Akbulut1,2, Cemil Colak1

  • 1Department of Biostatistics and Medical Informatics, Inonu University Faculty of Medicine, 44280 Malatya, Turkey.

PubMed

Insights

Machine learning models accurately classify pediatric acute appendicitis (AAP) subtypes. Random Forest and XGBoost show promise in improving diagnosis and patient outcomes by distinguishing between negative, uncomplicated, and complicated cases.

Area of Science:

  • Computational biology and bioinformatics
  • Pediatric surgery and emergency medicine
  • Artificial intelligence in healthcare

Background:

  • Pediatric acute appendicitis (AAP) diagnosis is challenging, with misclassification risking delayed treatment or unnecessary surgery.
  • Accurate classification into negative, uncomplicated, and complicated AAP is crucial for appropriate pediatric care.
  • Existing diagnostic methods for AAP require enhancement to improve precision and patient outcomes.

Purpose of the Study:

  • To evaluate and compare the diagnostic accuracy of five machine learning (ML) models for classifying pediatric AAP subtypes.
  • To identify the most effective ML models for distinguishing between negative, uncomplicated, and complicated pediatric appendicitis.
  • To assess the role of specific laboratory biomarkers in the ML-based classification of AAP.

Main Methods:

  • Retrospective analysis of 590 pediatric patients diagnosed with AAP.
  • Inclusion of demographic data and laboratory parameters (CRP, WBC, neutrophils, lymphocytes, appendiceal diameter) as features.
  • Training and testing of five ensemble ML models (AdaBoost, XGBoost, Stochastic Gradient Boosting, Bagged CART, Random Forest) using cross-validation.

Main Results:

  • Random Forest achieved 90.7% accuracy, 100% sensitivity, and 61.5% specificity for negative vs. uncomplicated AAP.
  • XGBoost demonstrated superior performance for complicated AAP with 97.3% accuracy, 100% sensitivity, and 78.3% specificity.
  • Neutrophil count, appendiceal diameter, and WBC levels were identified as the most influential predictive biomarkers.

Conclusions:

  • Machine learning models, specifically Random Forest and XGBoost, show significant potential in aiding pediatric AAP diagnosis.
  • ML-based decision support tools can enhance clinical judgment, leading to improved diagnostic accuracy and patient outcomes.
  • Future research should focus on multi-center validation, integrating imaging data, and improving model interpretability for clinical adoption.