PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 23, 2026GigaScience0 citationsOpen Access

A preregistered, open pipeline for early cerebral palsy risk assessment from Infant Videos

View Full Paper
MSMelanie SegadoUniversity of PennsylvaniaLPLaura A. ProsserChildren's Hospital of PhiladelphiaADAndrea F Duncan

Key Points

  • The aim is to develop a standardized approach for early assessment of cerebral palsy risk using infant videos and machine learning.
  • Developed an end-to-end machine learning pipeline with off-the-shelf pose estimation and general feature extraction.
  • Applied the pipeline to a large dataset of 1053 infants from a high-risk clinical cohort.
  • Used a 'lock-box' test set for unbiased model evaluation.
  • Achieved moderate predictive accuracy for GMA scores (ROC-AUC = 0.77).
  • Reflected good performance despite a low prevalence of adverse GMA outcomes (10-12%).
  • Provided de-identified feature data and open-source code for future research.

Abstract

Abstract Cerebral Palsy (CP), affecting approximately 1 in 500 children due to abnormal brain development, impacts movement control. Early risk assessment via the General Movements Assessment (GMA) at 3-4 months is highly predictive for CP but relies on trained clinicians. Machine-learning-based approaches for predicting GMA score from video have shown considerable promise, but typically rely on dataset-specific preprocessing, custom feature sets, and manually designed model pipelines, which make external benchmarking more difficult. This, combined with strict privacy constraints on sharing data, makes it challenging to train and evaluate models across datasets, which is important for assessing clinical utility. There is therefore a need to develop approaches that will work across different datasets to enable multi-site dataset aggregation and model training. To address this gap, we developed an end-to-end pipeline that uses off-the-shelf pose estimation, general-purpose feature extraction, and automated machine learning — none of which are tuned to a specific dataset. We applied this approach to a newly generated large dataset of 1053 infants (with approximately 10—12% positive class for adverse GMA outcome, drawn from a high-risk clinical cohort) within a preregistered study design. Model performance was evaluated on a strict “lock-box” test set, which remained untouched during any phase of model development or preprocessing optimization, and only used for evaluation once the final model and pipeline had been preregistered. The developed model achieved moderate predictive accuracy for clinician-assessed GMA scores (Area Under the Receiver Operating Characteristic Curve, ROC-AUC = 0.77; Area Under the Precision-Recall Curve, PR-AUC = 0.41). The moderate accuracy is noteworthy given the 10—12% positive class prevalence, and power-law scaling of ROC-AUC as a function of increasing dataset size. By releasing de-identified feature data and open-source code, and simplifying the training pipeline using AutoML, our work establishes essential groundwork for future robust, globally relevant CP screening tools suitable for low-resource settings.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Segado et al. (2026) studied this question.

synapsesocial.com/papers/69730f9fc8125b09b0d1f55chttps://doi.org/10.1093/gigascience/giag003
Ask AI
Helpful
Bookmark
Share
View Full Paper