Related Experiment Video
Updated: Jan 2, 2026

A Two-interval Forced-choice Task for Multisensory Comparisons
Published on: November 9, 2018
Bayesian analysis of paired-comparison sound quality ratings
Arne Leijon1, Martin Dahlquist2, Karolina Smeds2
1School of Electrical Engineering, KTH, Stockholm, Sweden.
Abstract:
This paper presents a method to analyze paired-comparison data including either binary or graded ordinal responses, with or without ties. The proposed method can use either of two classical choice models: (1) Thurstone case V, which assumes a Gaussian distribution of the sensory variables underlying listener decisions, or (2) the Bradley-Terry-Luce (BTL) model, which assumes a logistic distribution. The analysis method was validated using simulated paired-comparison experiments with known distributions of the sound-quality parameters in the simulated population from which "participants" were generated at random. The validation indicated that the Thurstone and BTL models give similar results close to the true values. The estimated credibility of a quality difference was slightly higher with the BTL model. The analysis results showed dramatically better precision when the response data included graded ordinal judgments instead of binary responses. Allowing tied responses also tended to improve precision. The method was also applied to data from a real evaluation of hearing-aid programs. The analysis revealed clinically interesting results with high statistical credibility, although the amount of test data was limited.
Related Concept Videos
Expected Frequencies in Goodness-of-Fit Tests
Friedman Two-way Analysis of Variance by Ranks
Wilcoxon Signed-Ranks Test for Matched Pairs
Bonferroni Test
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Multiple Comparison Tests
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...

