Related Experiment Video
Updated: Nov 10, 2025

Qualitative and Quantitative Validation of Tools with Rating Scales Aimed at Assessing the Quality of University Service-Learning
Published on: August 29, 2025
Beyond Likert ratings: Improving the robustness of developmental research measurement using best-worst scaling
Nichola Burton1, Michael Burton2, Carmen Fisher3
1ARC Center of Excellence in Cognition and its Disorders, School of Psychology, University of Western Australia, Crawley, Australia.
Abstract:
Some of the 'best practice' approaches to ensuring reproducibility of research can be difficult to implement in the developmental and clinical domains, where sample sizes and session lengths are constrained by the practicalities of recruitment and testing. For this reason, an important area of improvement to target is the reliability of measurement. Here we demonstrate that best-worst scaling (BWS) provides a superior alternative to Likert ratings for measuring children's subjective impressions. Seventy-three children aged 5-6 years rated the trustworthiness of faces using either Likert ratings or BWS over two sessions. Individual children's ratings in the BWS condition were significantly more consistent from session 1 to session 2 than those in the Likert condition, a finding we also replicate with a large adult sample (N = 72). BWS also produced more reliable ratings at the group level than Likert ratings in the child sample. These findings indicate that BWS is a developmentally appropriate response format that can deliver substantial improvements in reliability of measurement, which can increase our confidence in the robustness of findings with children.
More Related Videos
Related Concept Videos
Ordinal Level of Measurement
Data measured using an ordinal scale are similar to nominal scale data, but there is one major difference. The ordinal scale data can be ordered. An example of ordinal scale data is a list of the top five national parks...
Ratio Level of Measurement
A set of data measured using the ratio scale takes care of the ratio problem and provides complete information. Ratio scale data are like interval scale data, except they have a zero point and ratios can be calculated....
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Regression Toward the Mean
Self-Report Tests of Personality
Reliability and Validity

