Related Experiment Video
Updated: Jun 4, 2026

Assessment of Mouse Judgment Bias through an Olfactory Digging Task
Published on: March 4, 2022
A validation of Amazon Mechanical Turk for the collection of acceptability judgments in linguistic theory
1Department of Cognitive Sciences, University of California, 3151 Social Science Plaza A, Irvine, CA 92697-5100, USA. jsprouse@uci.edu
Abstract:
Amazon's Mechanical Turk (AMT) is a Web application that provides instant access to thousands of potential participants for survey-based psychology experiments, such as the acceptability judgment task used extensively in syntactic theory. Because AMT is a Web-based system, syntacticians may worry that the move out of the experimenter-controlled environment of the laboratory and onto the user-controlled environment of AMT could adversely affect the quality of the judgment data collected. This article reports a quantitative comparison of two identical acceptability judgment experiments, each with 176 participants (352 total): one conducted in the laboratory, and one conducted on AMT. Crucial indicators of data quality--such as participant rejection rates, statistical power, and the shape of the distributions of the judgments for each sentence type--are compared between the two samples. The results suggest that aside from slightly higher participant rejection rates, AMT data are almost indistinguishable from laboratory data.
Related Concept Videos
Hypothesis: Accept or Fail to Reject?
There are two ways to indicate that the null hypothesis is not rejected. 'Accept' the null hypothesis and 'fail to...
Stereotype Content Model
Lazarus's Cognitive Appraisal Theory
Primary Appraisal:...
Reliability and Validity
Stereotypes, Prejudice, and Discrimination