Related Experiment Video
Updated: May 3, 2026

A Modified Trier Social Stress Test for Vulnerable Mexican American Adolescents
Published on: July 10, 2017
Getting serious about test-retest reliability: a critique of retest research and some recommendations
1Humanalysis, Inc., Saratoga Springs, NY, USA, denisefpolit@gmail.com.
Purpose:
To focus attention on the need for rigorous and carefully designed test-retest reliability assessments for new patient-reported outcomes and to encourage retest researchers to be thoughtful, ambitious, and creative in their retest efforts.
Methods:
The paper outlines key challenges that confront retest researchers, calls attention to some limitations in meeting those challenges, and describes some strategies to improve retest research.
Results:
Modest retest coefficients are often reported as acceptable, and many important decisions-such as the retest interval-appear not to be evidence-based. Retest assessments are seldom undertaken before a measure has been finalized, which rules out using retest data to select strong, reproducible items.
Conclusions:
Strategies for improving retest research include seeking input from patients or experts regarding the stability of the construct to support decisions about the retest interval, analyzing item-level retest data to identify items to revise or discard, establishing a priori standards of acceptability for reliability coefficients, using large, heterogeneous, and representative retest samples and collecting follow-up data to better understand consistent and inconsistent responses over time.
More Related Videos
08:40Isokinetic Robotic Device to Improve Test-Retest and Inter-Rater Reliability for Stretch Reflex Measurements in Stroke Patients with Spasticity
Published on: June 12, 2019
08:06Testing for Metacognitive Responding Using an Odor-based Delayed Match-to-Sample Test in Rats
Published on: June 18, 2018
Related Concept Videos
Reliability and Validity
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
Comparing Experimental Results: Student's t-Test
Bioequivalence Experimental Study Designs: Repeated Measures, Cross-Over, Carry-Over, and Latin Square Designs
Introspection
Regression Toward the Mean