Related Experiment Video
Updated: Aug 9, 2025

An R-Based Landscape Validation of a Competing Risk Model
Published on: September 16, 2022
An integrated approach to geographic validation helped scrutinize prediction model performance and its variability
Tsvetan R Yordanov1, Ricardo R Lopes2, Anita C J Ravelli1
1Amsterdam UMC location University of Amsterdam, Medical Informatics, Meibergdreef 9, Amsterdam, The Netherlands; Amsterdam Public Health Research Institute, Amsterdam, The Netherlands.
Objectives:
To illustrate in-depth validation of prediction models developed on multicenter data.
Methods:
For each hospital in a multicenter registry, we evaluated predictive performance of a 30-day mortality prediction model for transcatheter aortic valve implantation (TAVI) using the Netherlands heart registration (NHR) dataset. We measured discrimination and calibration per hospital in a leave-center-out analysis (LCOA). Meta-analysis was used to calculate I2 values per performance metric from the LCOA and to compute mean and confidence interval (CI) estimates. Case mix differences between studies were inspected using the framework of Debray et al. for understanding external validation. We also aimed to discover subgroups (SGs) with high model prediction error (PE) and their distribution over the centers.
Results:
We studied 16 hospitals with 11,599 TAVI patients with an early mortality of 3.7%. The models' area under the curve (AUCs) had a wide range between hospitals from 0.59 to 0.79, and miscalibration occurred in seven hospitals. Mean AUC from meta-analysis was 0.68 (95% CI 0.65-0.70). I2 values were 0%, 74%, and 0% for AUC, calibration intercept and slope, respectively. Between-hospital case-mix differences were substantial, and model transportability was low. One SG was discovered with marked global PE and was associated with poor performance on validation centers.
Conclusion:
The illustrated combination of approaches provides useful insights to inspect multicenter-based prediction models, and it exposes their limitations in transportability and performance variability when applied to different populations.
More Related Videos
Related Concept Videos
Data Validation
Key parameters for method validation include:
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Variation
When independent and dependent variables are plotted on a scatter plot, the slope of a line is a value that describes the rate of change between the two...
Reliability and Validity
Variability: Analysis
The range is a simple measure of variability, indicating the difference between the highest and...
Random Error

