Related Experiment Video
Updated: Sep 18, 2026

Combining Behavioral Endocrinology and Experimental Economics: Testosterone and Social Decision Making
Published on: March 2, 2011
Measurement in the Ultimatum and Dictator Games: Construct validity, standardization and the problem of coordination
1Department of Logic, History and Philosophy of Science, UNED (Universidad Nacional de Educación a Distancia), Paseo de Senda del Rey 7, 28040, Madrid, Spain. mjbuedo@fsof.uned.es.
Abstract:
The ultimatum game (UG) and the dictator game (DG) are two of the most prominent tools in experimental and behavioural social sciences. Originally designed to study bargaining behaviour and decision-making in economics, these simple yet versatile scripts have come to be viewed by many as reliable instruments for measuring prosocial preferences, such as altruism, generosity, or sensitivity to social norms. Proponents of the measurement view have argued that the UG and DG serve as standardized "social thermometers," capable of revealing cultural and contextual differences in normative behaviour across diverse populations. For instance, the UG has been used to demonstrate the influence of institutional contexts and cultural norms on fairness expectations, while the DG has been described as a straightforward measure of altruism. In this paper, we critically examine the idea that the UG and DG function as robust measuring devices, raising key questions about their validity as tools of measurement. We argue that the constructs these games are purported to measure-ranging from prosociality to norm compliance-are often imprecisely defined and vary across studies, and this severely limits the sense in which the DG or the UG can be interpreted as measuring devices. By analysing the ontological and methodological challenges surrounding these games, we contend that the UG and DG are better understood as tools for exploratory experimentation rather than as reliable instruments for systematic measurement.
Related Concept Videos
Measures of Intelligence
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this; it...
Social Foundations of Self II: The Generalized Other
Reliability and Validity
Strategies of Self-Presentation III: Self-Monitoring
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Theory of Attribution II: Kelley's Covariation Theory
