Related Experiment Video
Updated: Jan 22, 2026

Enhancing the Development and Growth of Infant Cerebral Palsy Rats Using Selective Spinal Manipulations
Published on: February 2, 2024
A preregistered, open pipeline for early cerebral palsy risk assessment from infant videos
Melanie Segado1, Laura A Prosser2,3, Andrea F Duncan2,4
1Department of Bioengineering, University of Pennsylvania, 220 S 33rd Street, Philadelphia, PA 19104, USA.
Insights
This study developed an automated pipeline to predict Cerebral Palsy (CP) risk from infant movement videos, aiming for broader clinical utility. The model achieved moderate accuracy, paving the way for accessible CP screening tools.
Area of Science:
- Neurology
- Medical Imaging
- Machine Learning
Background:
- Cerebral Palsy (CP) affects 1 in 500 children, impacting motor control due to abnormal brain development.
- General Movements Assessment (GMA) at 3-4 months is a key predictor for CP, but requires specialized clinicians.
- Current machine learning (ML) models for GMA prediction are often dataset-specific, hindering external validation and multi-site collaboration.
Purpose of the Study:
- To develop a generalized, end-to-end ML pipeline for predicting GMA scores from infant videos.
- To enable multi-site dataset aggregation and model training, overcoming privacy constraints and dataset-specific limitations.
- To establish groundwork for robust, globally relevant CP screening tools, particularly for low-resource settings.
Main Methods:
- An end-to-end pipeline was created using off-the-shelf pose estimation, general feature extraction, and automated machine learning (AutoML).
- The pipeline was applied to a new dataset of 1053 infants from a high-risk cohort.
- Model performance was evaluated on a preregistered, untouched 'lock-box' test set.
Main Results:
- The model achieved moderate predictive accuracy for clinician-assessed GMA scores (ROC-AUC = 0.77, PR-AUC = 0.41).
- Performance is notable given the 10-12% positive class prevalence and known scaling properties of ROC-AUC.
- De-identified feature data and open-source code were released to facilitate future research.
Conclusions:
- The developed generalized pipeline addresses the need for cross-dataset compatible models in CP risk assessment.
- This approach simplifies model training and supports the development of more robust and accessible CP screening tools.
- The work contributes to advancing global efforts in early detection and management of Cerebral Palsy.
Abstract:
Cerebral palsy (CP), affecting approximately 1 in 500 children due to abnormal brain development, impacts movement control. Early risk assessment via the general movements assessment (GMA) at 3-4 months is highly predictive for CP but relies on trained clinicians. Machine-learning-based approaches for predicting GMA score from video have shown considerable promise, but typically rely on dataset-specific preprocessing, custom feature sets, and manually designed model pipelines, which make external benchmarking more difficult. This, combined with strict privacy constraints on sharing data, makes it challenging to train and evaluate models across datasets, which is important for assessing clinical utility. There is therefore a need to develop approaches that will work across different datasets to enable multi-site dataset aggregation and model training. To address this gap, we developed an end-to-end pipeline that uses off-the-shelf pose estimation, general-purpose feature extraction, and automated machine learning-none of which are tuned to a specific dataset. We applied this approach to a newly generated large dataset of 1053 infants (with approximately 10-12% positive class for adverse GMA outcome, drawn from a high-risk clinical cohort) within a preregistered study design. Model performance was evaluated on a strict "lock-box" test set, which remained untouched during any phase of model development or preprocessing optimization, and only used for evaluation once the final model and pipeline had been preregistered. The developed model achieved moderate predictive accuracy for clinician-assessed GMA scores (area under the receiver operating characteristic curve, ROC-AUC = 0.77; area under the precision-recall curve, PR-AUC = 0.41). The moderate accuracy is noteworthy given the 10-12% positive class prevalence, and power-law scaling of ROC-AUC as a function of increasing dataset size. By releasing de-identified feature data and open-source code, and simplifying the training pipeline using AutoML, our work establishes essential groundwork for future robust, globally relevant CP screening tools suitable for low-resource settings.
Related Concept Videos
Design Example: Analyzing Capacity Contours for Flood Risk Assessment
Relative Risk
Factors Affecting the Risk of Infection
The integrity and count of the white blood cells help the body resist pathogens and fight infection. When impaired, it reduces the body's resistance to pathogens. The acidic pH levels of the gastrointestinal, genitourinary tracts, and skin...
Drug Dosing: Infants and Children
Cerebral Hemispheres
Endoscopic Procedures III: Video Capsule Endoscopy

