Related Experiment Video
Updated: Jun 2, 2026

Adaptation of Semiautomated Circulating Tumor Cell (CTC) Assays for Clinical and Preclinical Research Applications
Published on: February 28, 2014
A log-adjusted t-statistic for large clinical laboratory datasets: a simulation study and real-world application
Mehmet Güven Günver1, Mehmet Fatih Sert2, Eray Yurtseven1
1Department of Biostatistics, Istanbul Faculty of Medicine, Istanbul University, Istanbul, Turkey.
Abstract:
In large datasets, conventional t-tests may identify statistically significant but practically trivial differences because statistical significance increases with sample size. A log-adjusted t-statistic, defined as an empirical sample-size-aware modification of the classical t-statistic, was evaluated to reduce this oversensitivity. Performance was assessed by Monte Carlo simulations of two-sample comparisons across sample sizes from 10 to 50,000 and effect sizes from δ = 0 to 1.0, and by application to a real clinical laboratory dataset comprising 464,145 participants. In simulations, the -adjusted statistic showed null rejection rates close to 0.05 across sample sizes, whereas the classical t-test became increasingly oversized at very large n. The adjustment was more conservative for small effects (δ = 0.2-0.4) while high rejection rates were retained for larger effects (δ = 0.6-1.0). In the real-data analysis, several sex differences that were highly significant by the classical t-test had small effect sizes and yielded reference p-values above the conventional 0.05 threshold after adjustment; platelet count (Cohen's d = 0.13) changed from p < 10-300 to reference p = 0.052, and potassium (d = 0.05) from p = 10-51 to reference p = 0.104. In contrast, larger effects such as hematocrit (d = 0.83) and HDL cholesterol (d = 0.77) continued to yield reference p-values below that threshold. These reference p-values were compared with the conventional α = 0.05 threshold for illustrative purposes only and were not intended to imply formal Type I error control. These findings suggest that the log-adjusted t-statistic may serve as a useful empirical decision aid for interpreting large clinical laboratory datasets by attenuating sample-size-driven significance while preserving detection of substantively meaningful effects.
Related Concept Videos
Comparing Experimental Results: Student's t-Test
Statistical Software for Data Analysis and Clinical Trials
Statistical Methods to Analyze Parametric Data: Student t-Test and Goodness-of-Fit Test
The Student's t-test is a statistical test that examines if there is a statistically significant difference between the means of two groups. This test is instrumental when dealing with data...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Significance Testing: Overview
Bioequivalence Data: Statistical Interpretation
