Related Experiment Video
Updated: Sep 7, 2025

08:05
Design and Analysis for Fall Detection System Simplification
Published on: April 6, 2020
10.8K
Assessing model accuracy using random data split: a simulation study
1Center for Device Evaluation and Radiological Health (CDRH), FDA, U.S. Food and Drug Administration, Silver Spring, Maryland, USA.
Journal of Biopharmaceutical Statistics
|June 22, 2022
Summary
Randomly splitting data for model testing limits assessing generalizability. True model performance evaluation requires distinct, independent datasets, especially when time drift is present.
Area of Science:
- Biostatistics
- Machine Learning in Healthcare
- Clinical Trial Methodology
Background:
- Randomization is a cornerstone in clinical studies to mitigate bias.
- Assessing model generalizability often involves splitting data into training and testing sets.
- The effectiveness of random data splitting for evaluating model generalizability requires scrutiny.
Purpose of the Study:
- To investigate the limitations of random data splitting in assessing model generalizability.
- To evaluate how random splits impact model performance metrics under different data conditions.
- To determine optimal strategies for testing model generalizability.
Main Methods:
- Conducted simulation studies using three distinct datasets (binary and continuous endpoints).
- Employed large sample sizes (n=10,000) for each simulation.
- Repeated random data splits (training/testing) 1,000 times per scenario to compare model performance.
Main Results:
- Random splits showed similar model performance (true/false positive fractions, mean-squared errors) between training and testing sets.
- Significant discrepancies emerged when time drift effects were present in the data.
- Model performance on randomly split data does not reliably indicate generalizability in the presence of temporal changes.
Conclusions:
- Random data splitting is insufficient for robustly assessing model generalizability, particularly with time-varying data.
- Evaluating model accuracy requires testing on data distinct from the training set.
- Independent, separate studies are recommended for validating model generalizability.
Related Concept Videos
Mechanistic Models: Compartment Models in Individual and Population Analysis
85
Mechanistic models are utilized in individual analysis using single-source data, but imperfections arise due to data collection errors, preventing perfect prediction of observed data. The mathematical equation involves known values (Xi), observed concentrations (Ci), measurement errors (εi), model parameters (ϕj), and the related function (ƒi) for i number of values. Different least-squares metrics quantify differences between predicted and observed values. The ordinary least...
85
Testing a Claim about Standard Deviation
2.5K
A complete procedure to test a claim about population standard deviation or population variance is explained here.
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
2.5K
Modeling and Similitude
325
Scaled modeling is a fundamental technique in engineering, enabling the study of large and complex systems by creating smaller, manageable replicas that recreate critical characteristics of the original. In hydrology and civil infrastructure, for example, scaled models of dams help analyze water flow, turbulence, and pressure. This method allows for accurate predictions of real-world behavior within a controlled environment, significantly reducing the cost and time involved in full-scale...
325
Typical Model Studies
436
Fluid mechanics model studies often utilize scaled-down systems to predict fluid behavior in full-scale environments, such as river flows, dam spillways, and structures interacting with open surfaces. Maintaining Froude number similarity in river models is crucial, as it replicates surface flow features like wave patterns and velocities.
436
Survival Tree
154
Survival trees are a non-parametric method used in survival analysis to model the relationship between a set of covariates and the time until an event of interest occurs, often referred to as the "time-to-event" or "survival time." This method is particularly useful when dealing with censored data, where the event has not occurred for some individuals by the end of the study period, or when the exact time of the event is unknown.
Building a Survival Tree
Constructing a...
Building a Survival Tree
Constructing a...
154
Random Sampling Method
12.1K
Sampling is a technique to select a portion (or subset) of the larger population and study that portion (the sample) to gain information about the population. Data are the result of sampling from a population. The sampling method ensures that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest. Among the various sampling methods used by...
12.1K

