Related Experiment Video
Updated: Jan 13, 2026

Implementation of a Real-Time Psychosis Risk Detection and Alerting System Based on Electronic Health Records using CogStack
Published on: May 15, 2020
Data-driven prediction of child neglect and abuse using integrated municipal sources
Naama Parush Shear Yashuv1, Rinat Salem2, Ofra Abramson2
1KI Research Institute, Kfar-Malal, Israel.
Background:
Child neglect and abuse are prevalent worldwide yet often incompletely reported and are frequently associated with long-term adverse physical and mental health outcomes. Municipal-level administrative data contain indicators relevant to detecting child neglect and abuse, which machine learning algorithms can aggregate to help identify children at-risk and facilitate timely interventions. However, this valuable information is typically stored in isolated data silos across different municipal services, limiting its effective utilization.
Objective:
This study aimed to assess whether machine learning models applied to integrated municipal data can accurately predict the risk of child neglect and abuse in a large population of children residing in Jerusalem, Israel.
Participants And Setting:
A large, deidentified dataset representing over 470,000 children, linked across multiple municipal systems, including population registry, education, public health, local taxation and welfare services.
Methods:
We defined neglect and abuse outcomes based on the child's welfare records, and constructed models to predict the current risk and the future 2-year risk for each outcome, using multitude of variables extracted from the dataset. Two main use cases were addressed: (1) risk prediction in the general child population using non-welfare data, and (2) risk prediction within the subpopulation already known to welfare services using both welfare and non-welfare data. The models were trained with incremental inclusion of data sources, and their performance was evaluated using the area under the receiver operating characteristic curve (AUC) and sensitivity at fixed levels of specificity.
Results:
The prediction models demonstrated good performance, with AUCs ranging from 0.75 to 0.88, depending on the use case and the time window for risk estimation. Accuracy improved with the integration of additional data sources, particularly education and taxation records. In a scenario where the top 5 % of children at risk, according to the algorithm, are assessed by municipal services, 32 % of neglected children and 34 % of abused children would have been identified up to 2 years in advance. Predictive performance was generally consistent across sex groups, but showed slightly lower AUCs for Arab children, compared to Jewish children.
Conclusions:
Machine learning models utilizing multi-source municipal data can effectively identify children at risk of maltreatment. Such tools may support municipal welfare systems by enhancing early detection, guiding resource allocation, and improving outcomes for vulnerable children. However, ethical considerations, cultural sensitivity, and human oversight are essential to ensure responsible implementation.
Related Concept Videos
Steps in Outbreak Investigation
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Applications of GIS: Disaster Management and Emergency Response
Levels of Use of a GIS

