Related Experiment Video
Updated: May 2, 2026

Implementation of a Real-Time Psychosis Risk Detection and Alerting System Based on Electronic Health Records using CogStack
Published on: May 15, 2020
Evaluating model generalizability for suicide attempt risk prediction: traditional machine vs deep learning
Nicholas Josselyn1,2, Sahil Sawant1,2,3, Rachel E Davis-Martin2
1Data Science, Worcester Polytechnic Institute, Worcester, MA, USA.
Abstract:
Suicide remains a leading cause of death and a significant public health concern in the United States. A majority (83%) of suicide decedents had a healthcare visit within the prior 365 days, presenting unique opportunities to utilize healthcare data for AI-based interventions. While previous works applied machine learning (ML) to analyze healthcare records for suicide attempt risk prediction (SARP), they lack external validation. Additionally, advantages of deep learning (DL) over ML for tabular SARP remains understudied. We performed external validation of a state-of-the-art SARP model from the Mental Health Research Network using over 750,000 UMass Memorial Health patient encounters. We further compared ML vs DL, assessing cross-setting healthcare generalizability. We found existing models did not generalize well, ML significantly outperformed DL on most metrics, and DL achieved higher sensitivity. These findings underscore the need for developing robust, generalizable SARP models for diverse healthcare contexts, improving identification of individuals at risk.