Related Experiment Video
Updated: Aug 18, 2026

Frailty Assessment in an Aging Mouse Model
Published on: September 23, 2025
Direct Comparison Between Loss-to-Follow-Up and Statistical Fragility Is Methodologically Inappropriate, and
Prushoth Vivekanantha1, Helena Son2, Jeffrey Kay1
1Division of Orthopaedic Surgery, Department of Surgery, McMaster University, Hamilton, Canada.
Purpose:
To assess the relation between the fragility index (FI) and reverse fragility index (RFI) with the minimum number of patients needed to reverse statistical significance (e.g. henceforth termed the lost to follow-up index (LTFI) and reverse LTFI (R-LTFI), respectively) and apply machine learning to identify which trial parameters are most important in FI, RFI, and continuous fragility index composition given the nonlinearity of these metrics.
Methods:
A total of 300,000 randomized controlled trials (100,000 for each metric) were simulated using common trial parameter value ranges. For FI and RFI, LTFI and R-LTFI values, respectively, were calculated as the minimum number of patients lost to follow-up to reverse significance in either direction. Machine learning models were trained to assess the relative importance of P value, sample size, and event numbers in FI and RFI composition and P value, group means, group standard deviations, and group sizes for continuous fragility index composition.
Results:
Among their respective cohorts of 100,000, the LTFI and R-LTFI were greater than FI and RFI in 84.9% and 95.5% of simulated randomized controlled trials, respectively. Random Forest and XGBoost machine learning models had near perfect accuracy in modeling fragility metrics (R2 ≥ 0.98), justifying its use over standard linear regression models in identifying trial parameters that are most important in determining fragility values. Feature importance analysis showed that the P value accounted for 79.8%, 71.9%, and 64.5% of variability in FI, RFI, and continuous fragility index values.
Conclusions:
Fragility metrics are primarily mathematical reflections of standard trial parameters and are heavily driven by the P value. Direct comparisons between fragility metrics and lost to follow-up are statistically inappropriate because LTFI and R-LTFI routinely exceed FI and RFI, respectively, and should be avoided in future fragility-based studies.
Clinical Relevance:
Understanding that statistical fragility metrics are mathematical transformations of trial parameters, and not independent measures of robustness, can prevent trial misinterpretation. Additionally, recognizing that the number of patients required to be lost to follow-up to reverse trial significance often exceeds fragility metrics should discourage inappropriate comparisons in future orthopaedic research.
Related Concept Videos
Assumptions of Survival Analysis
Comparing the Survival Analysis of Two or More Groups
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast, controlled...
Regression Toward the Mean
Bonferroni Test
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
