Related Experiment Video
Updated: Oct 11, 2025

Development of an Individual-Tree Basal Area Increment Model using a Linear Mixed-Effects Approach
Published on: July 3, 2020
Bayesian variable selection in hierarchical difference-in-differences models
James P Normington1, Eric F Lock1, Thomas A Murray1
1Division of Biostatistics, School of Public Health, 43353University of Minnesota, Minneapolis, MN, USA.
Abstract:
A popular method for estimating a causal treatment effect with observational data is the difference-in-differences model. In this work, we consider an extension of the classical difference-in-differences setting to the hierarchical context in which data cannot be matched at the most granular level. Our motivating example is an application to assess the impact of primary care redesign policy on diabetes outcomes in Minnesota, in which the policy is administered at the clinic level and individual outcomes are not matched from pre- to post-intervention. We propose a Bayesian hierarchical difference-in-differences model, which estimates the policy effect by regressing the treatment on a latent variable representing the mean change in group-level outcome. We present theoretical and empirical results showing a hierarchical difference-in-differences model that fails to adjust for a particular class of confounding variables, biases the policy effect estimate. Using a structured Bayesian spike-and-slab model that leverages the temporal structure of the difference-in-differences context, we propose and implement variable selection approaches that target sets of confounding variables leading to unbiased and efficient estimation of the policy effect. We evaluate the methods' properties through simulation, and we use them to assess the impact of primary care redesign of clinics in Minnesota on the management of diabetes outcomes from 2008 to 2017.
Related Concept Videos
Comparing the Survival Analysis of Two or More Groups
Friedman Two-way Analysis of Variance by Ranks
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
Parametric Survival Analysis: Weibull and Exponential Methods
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...
Distributions to Estimate Population Parameter
Survival Tree
Building a Survival Tree
Constructing a...

