Related Experiment Video
Updated: Mar 21, 2026

A Behavioral Test Battery for the Repeated Assessment of Motor Skills, Mood, and Cognition in Mice
Published on: March 2, 2019
Normative comparisons for large neuropsychological test batteries: User-friendly and sensitive solutions to minimize
Hilde M Huizenga1,2,3, Joost A Agelink van Rentergem1, Raoul P P P Grasman1,2
1a Department of Psychology , University of Amsterdam , Amsterdam , The Netherlands.
Introduction:
In neuropsychological research and clinical practice, a large battery of tests is often administered to determine whether an individual deviates from the norm. We formulate three criteria for such large battery normative comparisons. First, familywise false-positive error rate (i.e., the complement of specificity) should be controlled at, or below, a prespecified level. Second, sensitivity to detect genuine deviations from the norm should be high. Third, the comparisons should be easy enough for routine application, not only in research, but also in clinical practice. Here we show that these criteria are satisfied for current procedures used to assess an overall deviation from the norm-that is, a deviation given all test results. However, we also show that these criteria are not satisfied for current procedures used to assess test-specific deviations, which are required, for example, to investigate dissociations in a test profile. We therefore propose several new procedures to assess such test-specific deviations. These new procedures are expected to satisfy all three criteria.
Method:
In Monte Carlo simulations and in an applied example pertaining to Parkinson disease, we compare current procedures to assess test-specific deviations (uncorrected and Bonferroni normative comparisons) to new procedures (Holm, one-step resampling, and step-down resampling normative comparisons).
Results:
The new procedures are shown to: (a) control familywise false-positive error rate, whereas uncorrected comparisons do not; (b) have higher sensitivity than Bonferroni corrected comparisons, where especially step-down resampling is favorable in this respect; (c) be user-friendly as they are implemented in a user-friendly normative comparisons website, and as the required normative data are provided by a database.
Conclusion:
These new normative comparisons procedures, especially step-down resampling, are valuable additional tools to assess test-specific deviations from the norm in large test batteries.
More Related Videos
06:23The 4 Mountains Test: A Short Test of Spatial Memory with High Sensitivity for the Diagnosis of Pre-dementia Alzheimer's Disease
Published on: October 13, 2016
07:02A Computerized Test Battery to Study Pharmacodynamic Effects on the Central Nervous System of Cholinergic Drugs in Early Phase Drug Development
Published on: February 11, 2019
Related Concept Videos
Self-Report Tests of Personality
Wechsler's Contribution to Measures of Intelligence
Measures of Intelligence
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Multiple Comparison Tests
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
Binet's Contribution to Measures of Intelligence