Related Experiment Video
Updated: Aug 22, 2025

Combining Multiple Data Acquisition Systems to Study Corticospinal Output and Multi-segment Biomechanics
Published on: January 9, 2016
Too Good to Be True: Bots and Bad Data From Mechanical Turk
Margaret A Webb1,2, June P Tangney2
1Department of Criminology, Max-Planck Institute for the Study of Crime, Security, and Law Freiburg im Breisgau.
Abstract:
Psychology is moving increasingly toward digital sources of data, with Amazon's Mechanical Turk (MTurk) at the forefront of that charge. In 2015, up to an estimated 45% of articles published in the top behavioral and social science journals included at least one study conducted on MTurk. In this article, I summarize my own experience with MTurk and how I deduced that my sample was-at best-only 2.6% valid, by my estimate. I share these results as a warning and call for caution. Recently, I conducted an online study via Amazon's MTurk, eager and excited to collect my own data for the first time as a doctoral student. What resulted has prompted me to write this as a warning: it is indeed too good to be true. This is a summary of how I determined that, at best, I had gathered valid data from 14 human beings-2.6% of my participant sample (N = 529).
Insights
Researchers caution against using Amazon
Area of Science:
- Behavioral and social sciences research
- Online data collection methodologies
Background:
- Psychology increasingly relies on digital data sources.
- Amazon's Mechanical Turk (MTurk) is a prominent platform for online studies.
- In 2015, an estimated 45% of top behavioral and social science journals featured MTurk studies.
Purpose of the Study:
- To report on the validity of data collected via Amazon's Mechanical Turk (MTurk).
- To serve as a cautionary example for researchers utilizing MTurk for data collection.
Main Methods:
- An online study was conducted using Amazon's Mechanical Turk (MTurk).
- Data validity was assessed to determine the proportion of reliable participant responses.
- Participant sample size was N = 529.
Main Results:
- An estimated 2.6% of the participant sample (14 out of 529 individuals) provided valid data.
- The overall data validity was significantly lower than anticipated.
Conclusions:
- The findings suggest that data collected through MTurk may not always be reliable.
- Researchers are urged to exercise caution and implement rigorous validation checks when using MTurk for data collection.
- The platform's efficiency in data gathering may be offset by concerns regarding data quality.
More Related Videos
Related Concept Videos
Regression Toward the Mean
Mechanical Efficiency of Real Machines
However, in reality, no machine can be truly ideal, and all of them experience some...
Stereotype Content Model

