Related Experiment Video
Updated: May 12, 2026

10:58
Multimedia Battery for Assessment of Cognitive and Basic Skills in Mathematics (BM-PROMA)
Published on: August 28, 2021
Assessing inconsistent responding in digital and paper formats: a replication study using the TIMSS 2023 Confidence
Evi Konstantinidou1, Vasiliki Pitsia2, Aidan Clerkin2
1Department of Psychology, University of Cyprus, Nicosia, Cyprus.
Frontiers in Psychology
|May 11, 2026
Summary
Digital testing shows higher inconsistent responding rates than paper formats, particularly in younger students. Mathematics achievement is a key factor in reducing this response inconsistency in large-scale assessments.
Area of Science:
- Educational measurement
- Psychometrics
- Survey methodology
Background:
- Inconsistent responding in surveys can bias results.
- Mixed-worded scales present unique challenges for respondents.
- Digital administration is increasingly common in large-scale assessments.
Purpose of the Study:
- To examine inconsistent responding in the TIMSS 2023 Confidence in Mathematics Scale.
- To compare digital versus paper questionnaire administrations.
- To identify predictors of inconsistent responding.
Main Methods:
- Posttest-only experimental design comparing digital and paper formats.
- Utilized Mean Absolute Difference and Factor Mixture Analysis for detection.
- Employed binary logistic regression to analyze student characteristics.
Main Results:
- Digital administration showed slightly higher inconsistent responding rates than paper.
- Fourth-grade students had higher rates of inconsistent responding than eighth-grade students.
- Mathematics achievement was the strongest negative predictor of inconsistent responding.
Conclusions:
- Administration mode impacts response consistency in large-scale assessments.
- Digital testing requires careful consideration to mitigate response bias.
- Understanding student characteristics, like math achievement, is crucial for improving survey data quality.
Related Concept Videos
Reliability and Validity
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
Measures of Intelligence
Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this; it...
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this; it...

