Related Experiment Video
Updated: Jun 30, 2025

Establishment of Rat Models Mimicking Gender-affirming Hormone Therapies
Published on: January 10, 2025
Comparing type 1 and type 2 error rates of different tests for heterogeneous treatment effects
Steffen Nestler1, Marie Salditt2
1University of Münster, Institut für Psychologie, Fliednerstr. 21, 48149, Münster, Germany. steffen.nestler@uni-muenster.de.
Abstract:
Psychologists are increasingly interested in whether treatment effects vary in randomized controlled trials. A number of tests have been proposed in the causal inference literature to test for such heterogeneity, which differ in the sample statistic they use (either using the variance terms of the experimental and control group, their empirical distribution functions, or specific quantiles), and in whether they make distributional assumptions or are based on a Fisher randomization procedure. In this manuscript, we present the results of a simulation study in which we examine the performance of the different tests while varying the amount of treatment effect heterogeneity, the type of underlying distribution, the sample size, and whether an additional covariate is considered. Altogether, our results suggest that researchers should use a randomization test to optimally control for type 1 errors. Furthermore, all tests studied are associated with low power in case of small and moderate samples even when the heterogeneity of the treatment effect is substantial. This suggests that current tests for treatment effect heterogeneity require much larger samples than those collected in current research.
Related Concept Videos
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
Errors In Hypothesis Tests
Test for Homogeneity
Bonferroni Test
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
Types of Hypothesis Testing
When the null and alternative hypotheses are stated, it is observed that the null hypothesis is a neutral statement against which the alternative hypothesis is tested. The alternative hypothesis is a claim that instead has a certain direction. If the null hypothesis claims that p = 0.5, the alternative hypothesis would be an opposing statement to this and can be put either p > 0.5, p < 0.5, or p...
Comparing Experimental Results: Student's t-Test

