Multilevel quantile function modeling with application to birth outcomes

Luke B Smith1, Brian J Reich1, Amy H Herring2

  • 1Department of Statistics, North Carolina State University, Raleigh, North Carolina 27695-8203, U.S.A.

Biometrics
|March 13, 2015
PubMed

Insights

Air pollution, specifically ozone, is linked to adverse birth outcomes like lower gestational age and birth weight in Texas infants. This study introduces a novel Bayesian approach to analyze these complex environmental health relationships.

Area of Science:

  • Environmental epidemiology
  • Biostatistics
  • Perinatal health

Background:

  • Infants born preterm or small for gestational age face higher morbidity and mortality risks.
  • Understanding environmental factors influencing birth outcomes is crucial for public health.

Purpose of the Study:

  • To investigate the association between ozone exposure and birth weight and gestational age in Texas infants.
  • To develop and apply a flexible semi-parametric Bayesian quantile model for analyzing environmental health data.

Main Methods:

  • Utilized Texas birth certificate data (2002-2004) and EPA air pollution estimates.
  • Employed a semi-parametric Bayesian multilevel quantile function model to analyze the full distribution of birth weight and gestational age.
  • Incorporated extreme value theory for low birth weight analysis and methods for discrete response data.

Main Results:

  • Ozone exposure was negatively associated with the lower tail of gestational age in South Texas.
  • Ozone exposure showed a negative association with the distribution of birth weight for high gestational ages.
  • The proposed modeling approach demonstrated reduced mean squared error in effect estimation through information pooling.

Conclusions:

  • Environmental factors like ozone can significantly impact infant birth outcomes, particularly at the extremes of the distribution.
  • The developed Bayesian quantile methodology provides a robust framework for analyzing complex environmental health associations.
  • The R package BSquare offers accessible tools for implementing these advanced statistical methods.

Related Concept Videos

Parametric Survival Analysis: Weibull and Exponential Methods01:14

Parametric Survival Analysis: Weibull and Exponential Methods

Parametric survival analysis models survival data by assuming a specific probability distribution for the time until an event occurs. The Weibull and exponential distributions are two of the most commonly used methods in this context, due to their versatility and relatively straightforward application.
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...
1.3K
Applications of Life Tables01:22

Applications of Life Tables

Life tables are versatile across various fields, providing a quantitative basis for analyzing mortality and survival rates. Whether used by demographers, actuaries, epidemiologists, or sociologists, life tables offer valuable insights into the dynamics of life and death, facilitating informed decisions in public health, insurance, conservation, and beyond. Their broad applicability highlights the interconnectedness of demographic data with practical outcomes in everyday life and strategic...
423
z Scores and Area Under the Curve01:17

z Scores and Area Under the Curve

z scores are the standardized values obtained after converting a normal distribution into a standard normal distribution. A z score is measured in units of the standard deviation. The z score tells you how many standard deviations the value x is above (to the right of) or below (to the left of) the mean, μ. Values of x that are larger than the mean have positive z scores, and values of x that are smaller than the mean have negative z scores. If x equals the mean, then x has a z score of...
20.3K
Distributions to Estimate Population Parameter01:26

Distributions to Estimate Population Parameter

The accurate values of population parameters such as population proportion, population mean, and population standard deviation (or variance) are usually unknown. These are fixed values that can only be estimated from the data collected from the samples. The estimates of each of these parameters are sample proportion, the sample mean, and sample standard deviation (or variance). To obtain the values of these sample statistics, data are required that have particular distribution and central...
5.8K
Quartile01:15

Quartile

Quartiles are numbers that separate the data into quarters. Quartiles may or may not be part of the data. To find the quartiles, first, find the median or second quartile. The first quartile, Q1, is the middle value of the lower half of the data, and the third quartile, Q3, is the middle value, or median, of the upper half of the data. To get the idea, consider the same data set:
1; 1; 2; 2; 4; 6; 6.8; 7.2; 8; 8.3; 9; 10; 10; 11.5
The median or second quartile is seven. The lower half of the...
10.1K
Truncation in Survival Analysis01:09

Truncation in Survival Analysis

Truncation in survival analysis refers to the exclusion of individuals or events from the dataset based on specific criteria related to the time of the event. This exclusion can happen in two primary forms: left truncation and right truncation.
Left truncation occurs when individuals who experienced the event of interest before a certain time are not included in the study. This is often due to a "delayed entry" into the study where only those who survive until a certain entry point are...
725