Related Experiment Video
Updated: Jan 9, 2026

Assessment of Stress Effects on Cognitive Flexibility using an Operant Strategy Shifting Paradigm
Published on: May 4, 2020
Do sex differences influence test habituation and internal data validity in neurocognitive testing? A blinded
Konstantin Warneke1, Manuel Oraze2, Marco Herbsleb3
1Institute for Sustainability Psychology, Leuphana University Lüneburg, Lüneburg, Germany; Institute of Sport Science, Department for Human Movement Science and Exercise Physiology, Friedrich Schiller University Jena, Jena, Germany.
Abstract:
Reliable neurocognitive assessment requires sufficient habituation to ensure that test outcomes reflect stable cognitive performance rather than learning effects. This study examined the influence of repeated testing and sex differences on the reliability and internal validity of three widely used neurocognitive tasks: the Trail Making Test, Stroop Test (Word Read and Color Read), and CRT. One hundred healthy young adults (47 men, 53 women) completed all tasks twice daily over five consecutive days. Relative and absolute reliability, as well as agreement metrics were calculated to quantify systematic and random errors. Significant within- and between-days habituation effects were observed. Reliability varied substantially: the reaction tasks showed the highest stability, followed by Stroop tasks; the Trail-Making-Test B demonstrated the lowest reproducibility. Systematic improvements were most pronounced between sessions one and two and generally stabilized after two to four days of familiarization. Sex-specific analyses revealed consistent male superiority in choice reaction performance. Sex differences in habituation were task-dependent and primarily reflected differences in adaptation rate rather than the magnitude of improvement. Across sexes, sufficient task familiarization was essential to minimize systematic and random errors. Overall reliability metrics were similar across sexes. Maximal random errors were reported in the Trail-Making-Test, contradicting unhabituated test application to track longitudinal changes or establishing valid cross-sectional analyses.
More Related Videos
05:53A Metric Test for Assessing Spatial Working Memory in Adult Rats Following Traumatic Brain Injury
Published on: May 7, 2021
10:02Assessment of Spontaneous Alternation, Novel Object Recognition and Limb Clasping in Transgenic Mouse Models of Amyloid-β and Tau Neuropathology
Published on: May 28, 2017
Related Concept Videos
Blind Procedures
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
Sign Test for Matched Pairs
To conduct the sign test, we first calculate the differences in...