Related Experiment Video
Updated: Mar 16, 2026

Measuring Cardiac Autonomic Nervous System ANS Activity in Toddlers - Resting and Developmental Challenges
Published on: February 25, 2016
Analysis of multiple-variable missing-not-at-random survey data for child lead surveillance using NHANES
Eric M Roberts1, Paul B English2
1Public Health Institute, Oakland, CA, 94607, U.S.A.
A new Bayesian model addresses missing data in health surveys, improving surveillance of elevated blood lead levels (EBLLs) in children. This method reveals significant disparities in EBLLs across demographics and housing, enhancing public health insights.
Area of Science:
- Public Health Surveillance
- Biostatistics
- Environmental Health
Background:
- Multi-topic surveys are crucial for public health surveillance but often suffer from high proportions of missing data.
- The National Health and Examination Survey (NHANES) is vital for tracking elevated blood lead levels (EBLLs) in US children, yet up to 35% of respondents have missing key predictor variables.
Purpose of the Study:
- To develop and apply a statistical model to address missing-not-at-random variables in complex survey data for improved public health surveillance.
- To estimate the prevalence of elevated blood lead levels in young children using a novel Bayesian approach applied to the American Community Survey (ACS).
Main Methods:
- Formulated a t-distributed Heckman selection model within a Bayesian framework to handle multiple missing-not-at-random variables in complex survey designs.
- Utilized Gibbs and grid sampling to estimate posterior distributions of parameters.
- Applied the developed model coefficients to ACS data to calculate prevalence estimates for EBLLs.
Main Results:
- Quantified significant disparities in EBLL prevalence based on race/ethnicity, age of housing, and poverty.
- Presented three- to five-fold differences in predicted EBLL prevalence across various US geographies.
- Enabled multivariate analyses of EBLLs, including the critical variable of housing age, which was previously unavailable.
Conclusions:
- The developed Bayesian model enhances the utility of existing survey data for public health surveillance, particularly for tracking elevated blood lead levels.
- The findings highlight significant socioeconomic and environmental factors influencing EBLLs in children.
- This methodology offers a valuable expansion for NHANES and similar health surveillance programs facing challenges with missing data.
Related Concept Videos
Data Collection by Observations
An astronomer viewing the motion and brightness of stars in the sky and recording the data is an example of observational data collection. A botanist recording...
Comparing the Survival Analysis of Two or More Groups
Censoring Survival Data
Longitudinal Studies
Observational Studies
There are three types of observational studies – Prospective, retrospective, and cross-sectional.
Prospective Study
Prospective studies, also known as longitudinal or cohort studies, are carried out by collecting future data from groups sharing similar characteristics. One...
Longitudinal Research

