Related Experiment Video
Updated: Jul 23, 2025

Implementation of a Real-Time Psychosis Risk Detection and Alerting System Based on Electronic Health Records using CogStack
Published on: May 15, 2020
Validation of a Bayesian learning model to predict the risk for cannabis use disorder
Thanthirige Lakshika M Ruberu1, Rajapaksha Mudalige Dhanushka S Rajapaksha1, Mary M Heitzeg2
1Department of Mathematical Sciences, University of Texas at Dallas, Richardson, TX 75080, United States.
Background:
Cannabis use disorder (CUD) is a growing public health problem. Early identification of adolescents and young adults at risk of developing CUD in the future may help stem this trend. A logistic regression model fitted using a Bayesian learning approach was developed recently to predict the risk of future CUD based on seven risk factors in adolescence and youth. A nationally representative longitudinal dataset, Add Health was used to train the model (henceforth referred as Add Health model).
Methods:
We validated the Add Health model on two cohorts, namely, Michigan Longitudinal Study (MLS) and Christchurch Health and Development Study (CHDS) using longitudinal data from participants until they were approximately 30 years old (to be consistent with the training data from Add Health). If a participant was diagnosed with CUD at any age during this period, they were considered a case. We calculated the area under the curve (AUC) and the ratio of expected and observed number of cases (E/O). We also explored recalibrating the model to account for differences in population prevalence.
Results:
The cohort sizes used for validation were 424 (53 cases) for MLS and 637 (105 cases) for CHDS. AUCs for the two cohorts were 0.66 (MLS) and 0.73 (CHDS) and the corresponding E/O ratios (after recalibration) were 0.995 and 0.999.
Conclusion:
The external validation of the Add Health model on two different cohorts lends confidence to the model's ability to identify adolescent or young adult cannabis users at high risk of developing CUD in later life.
Related Concept Videos
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Model Approaches for Pharmacokinetic Data: Distributed Parameter Models
The distributed parameter models are specifically designed to account for variations and differences in some drug classes. This model is particularly useful for assessing regional concentrations of anticancer or...
Pharmacokinetic Models: Comparison and Selection Criterion
Physiological models take a detailed approach by considering specific molecular processes. They can predict drug distribution, metabolism, and elimination changes, providing a comprehensive understanding of how drugs interact with the body.
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
Cancer Survival Analysis

