Related Experiment Video
Updated: May 2, 2026

An R-Based Landscape Validation of a Competing Risk Model
Published on: September 16, 2022
Classifier calibration using splined empirical probabilities in clinical risk prediction
René Gaudoin1, Giovanni Montana, Simon Jones
1Imperial College London, London, UK, r.gaudoin@imperial.ac.uk.
This study introduces splined empirical probabilities, a novel machine learning method for accurate probability estimation. It outperforms standard calibration techniques on diverse datasets, enhancing classification and risk assessment.
Area of Science:
- Machine Learning
- Statistical Modeling
- Data Science
Background:
- Supervised machine learning (ML) applications primarily focus on classification and ranking.
- Accurate probability estimation is crucial for many ML applications but remains a challenge.
- Existing methods for converting ML rankings into probability estimates have limitations.
Purpose of the Study:
- To present a novel method, splined empirical probabilities, for deriving accurate probability estimates from ML outputs.
- To compare the performance of splined empirical probabilities against existing calibration methods.
- To demonstrate the method's utility on simulated and real-world healthcare data.
Main Methods:
- Developed a new probability estimation technique based on the receiver operating characteristic (ROC) curve.
- Utilized splined empirical probabilities, a cumulative approach, integrated with ROC analysis.
- Evaluated performance using metrics like Hosmer-Lemeshow, Kullback-Leibler divergence, and cumulative distribution function differences.
Main Results:
- Splined empirical probabilities demonstrated favorable performance compared to the standard pool adjacent violators algorithm for isotonic regression.
- The method showed effectiveness across various measures of probability estimate quality.
- Successful application on both simulated and real healthcare datasets was observed.
Conclusions:
- Splined empirical probabilities offer a valuable and easily integrated alternative for accurate probability estimation in ML.
- The method provides a robust approach for enhancing ML model calibration.
- This technique holds significant potential for applications requiring precise probability assessments, particularly in healthcare.
More Related Videos
07:31Implementation of a Real-Time Psychosis Risk Detection and Alerting System Based on Electronic Health Records using CogStack
Published on: May 15, 2020
06:46Competing-Risk Nomogram for Predicting Cancer-Specific Survival in Multiple Primary Colorectal Cancer Patients after Surgery
Published on: September 27, 2024
Related Concept Videos
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
Calibration Curves: Correlation Coefficient
Calibration Curves: Linear Least Squares
For data that follow a straight line, the standard method for fitting is the linear...
Relative Risk
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.