Related Experiment Video
Updated: Sep 19, 2025

Author Spotlight: A Novel Setup to Conduct Naturalistic Laboratory Experiments with Real Human Actors in Scenarios
Published on: August 4, 2023
Regression-based normative data in neuropsychology: Using raw scores as observed response variable outperforms
Javier Oltra-Cucarella1, Rubén Pérez-Elvira2, Beatriz Bonete-López1
1Department of Health Psychology, Universidad Miguel Hernandez de Elche.
Abstract:
Regression-based normative data for neuropsychological variables are increasing in popularity over the last years. However, some use raw data while others use transformation when the observed response variable is skewed. This work analyzes how well the linear models fit for each type of variable. We used real data from a sample of n = 163 cognitively healthy individuals and compared the fit of linear regression models for raw scores and for corrected scaled scores. We then simulated a population of 1,000,000 individuals and drew 1,000 random samples of different sizes (n = 100, 200, 5,000, 1,000, 10,000) for seven different scenarios, analyzed the percentage of individuals scoring in the lowest 5%, and analyzed the agreement between models with the Cohen's κ statistic. Linear models for raw scores and for scaled scores were similar when the model included all the covariates, but barely identified low scores when scaled scores were corrected with covariates taken from different regressions (κ = 0.58). Models with raw scores showed that the expected number of individuals scoring low was close to the expected 5%, whereas models with scaled scores with covariates taken from different regressions were close to 0%. The two models agreed only when the response variable was random symmetrical and uncorrelated with the covariates. When calculating normative data using linear regressions, raw scores should be the preferred choice. If residuals analysis shows that the model does not fit the data well, researchers should consider using nonlinear models. Transforming data for normality of the observed response is discouraged. (PsycInfo Database Record (c) 2025 APA, all rights reserved).
More Related Videos
08:05Measuring Statistical Learning Across Modalities and Domains in School-Aged Children Via an Online Platform and Neuroimaging Techniques
Published on: June 30, 2020
06:49A Quick Phenotypic Neurological Scoring System for Evaluating Disease Progression in the SOD1-G93A Mouse Model of ALS
Published on: October 6, 2015
Related Concept Videos
Regression Toward the Mean
Introduction to Nonparametric Statistics
One of...
Central Limit Theorem
The sample size, n, that...
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
Detection of Gross Error: The Q Test
Testing a Claim about Mean: Known Population SD
Estimating a population mean requires the samples to be distributed normally. The data should be collected from the randomly selected samples having no sampling bias. The sample size needed to be higher than 30, and most importantly, the population standard deviation should be already known.
In most realistic situations, the population standard deviation is often unknown, but in rare circumstances, when it...