Related Experiment Video
Updated: May 15, 2025

Measuring the Structure, Composition, and Change of Underwater Environments with Large-area Imaging
Published on: April 18, 2025
AI-imputed and crowdsourced price data show strong agreement with traditional price surveys in data-scarce
Julius Adewopo1,2, Bo Pieter Johannes Andrée1, Helen Peter2
1Development Data Group, World Bank, Washington, DC, United States of America.
Abstract:
Continuous access to up-to-date food price data is crucial for monitoring food security and responding swiftly to emerging risks. However, in many food-insecure countries, price data is often delayed, lacks spatial detail, or is unavailable during crises when markets may become inaccessible, and rising prices can rapidly exacerbate hunger. Recent innovations, such as AI-driven data imputation and crowdsourcing, present new opportunities to generate continuous, localized price data. This paper evaluates the reliability of these approaches by comparing them to traditional enumerator-led data collection in northern Nigeria, a region affected by conflict, food insecurity, and data scarcity. The analysis examines crowdsourced prices for two staple food commodities, maize and rice, submitted daily by volunteers through a smartphone application over 36 months (2019-2021), and compares them with data collected concurrently by trained enumerators during the final eight months of 2021. Additionally, the crowdsourced dataset is compared to AI-imputed prices from the World Bank's Real-Time Prices (RTP) database. Data from the alternative methods reflected similar price inflation trends during the COVID-19 pandemic. Pearson's correlation coefficients indicate strong statistical agreement between crowdsourced and enumerator-collected prices (r = 0.94 for yellow and white maize, r = 0.96 for Indian rice, and r = 0.78 for Thailand rice). Furthermore, the crowdsourced data shows a high correlation with the AI-imputed prices (r = 0.99 for maize, and r = 0.94 for rice). The results from additional statistical tests of normality and paired means shows that the discrepancies between price datasets are consistent with measurement error rather than differences in actual price dynamics. Further tests of equivalence confirmed that enumerator and crowdsourced prices represent the same underlying market processes for specific commodity subtypes, and connotes that crowdsourced price data is a credible reference for validating AI-imputed estimates. The results support the use of AI imputation and crowdsourcing methods to improve price data collection and track market dynamics in near real time. These data innovations can be particularly valuable in areas that are underrepresented in national aggregate data due to limited monitoring capacity, and where high-frequency local data can aid targeted interventions.
More Related Videos
06:05The Participant-Reported Implementation Update and Score PRIUS: A Novel Method for Capturing Implementation-Related Data Over Time
Published on: February 19, 2021
06:16Signal Acquisition, Score Interpretation, and Economics of a Non-Invasive Point-of-Care Test for Coronary Artery Disease
Published on: August 9, 2024
Related Concept Videos
Data Collection by Survey
Data Collection by Observations
An astronomer viewing the motion and brightness of stars in the sky and recording the data is an example of observational data collection. A botanist recording...
Estimating Population Mean with Unknown Standard Deviation
William S. Gosset (1876–1937) of the...
Estimating Population Standard Deviation
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Stratified Sampling Method
To choose a stratified sample, divide the population into groups called strata and then take a...