Related Experiment Video
Updated: May 5, 2026

Multimedia Battery for Assessment of Cognitive and Basic Skills in Mathematics BM-PROMA
Published on: August 28, 2021
Missing data in a multi-item instrument were best handled by multiple imputation at the item score level
Iris Eekhout1, Henrica C W de Vet2, Jos W R Twisk1
1Department of Epidemiology and Biostatistics, VU University Medical Center, P.O. box 7057, 1007 MB Amsterdam, The Netherlands; EMGO Institute for Health and Care Research, VU University Medical Center, Van der Boechorststraat 7, 1081 BT Amsterdam, The Netherlands; Department of Methodology and Applied Biostatistics, Faculty of Earth and Life Sciences, Institute for Health Sciences, VU University, De Boelelaan 1085, 1081 HV Amsterdam, The Netherlands.
Multiple imputation (MI) for missing data in multi-item instruments provides more accurate regression estimates than mean imputation. Avoid mean imputation, especially with over 10% missing data, for reliable results.
Area of Science:
- Statistics
- Biostatistics
- Psychometrics
Background:
- Complete-case analysis is common despite available advanced methods like multiple imputation (MI).
- Handling missing data in multi-item instruments is crucial for accurate statistical modeling.
- The performance of various missing data handling techniques needs thorough evaluation.
Purpose of the Study:
- To compare the performance of simple and advanced missing data handling methods.
- To evaluate techniques for missing item scores in multi-item instruments.
- To identify optimal strategies for different missing data scenarios.
Main Methods:
- Simulated real-life missing data situations in a multi-item covariate for linear regression.
- Applied various missing data mechanisms with increasing percentages of missingness.
- Compared fitted regression coefficients using bias and coverage as performance metrics.
Main Results:
- Mean imputation resulted in biased estimates when over 10% of data were missing.
- Multiple imputation (MI) applied to individual item scores outperformed total score methods when >25% of subjects had missing items.
- MI demonstrated superior performance in handling substantial missing data.
Conclusions:
- Recommend using multiple imputation (MI) on item scores for accurate regression model estimates.
- Advise against using any form of mean imputation for handling missing data.
- Multiple imputation is the preferred method for missing data in multi-item instruments.
Related Concept Videos
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
One-Way ANOVA: Unequal Sample Sizes
Uncertainty in Measurement: Reading Instruments
One-Way ANOVA: Equal Sample Sizes
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Quantifying and Rejecting Outliers: The Grubbs Test
z Scores and Unusual Values
This score indicates how far a value is from the mean in terms of standard deviation. For example, if a data value has a z score of +1, the researcher can infer that the particular data value is one standard deviation above the mean. If another data...

