Related Experiment Video
Updated: Aug 29, 2026

Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations
Published on: February 15, 2017
Deviations from the population-averaged versus cluster-specific relationship for clustered binary data
Thomas R Ten Have1, Sarah J Ratcliffe, Beth A Reboussin
1Department of Biostatistics and Epidemiology, University of Pennsylvania School of Medicine, Blockley Hall, 6th FLR, 423 Guardian Drive, Philadelphia, PA 19104-6021, USA. ttenhave@cceb.upenn.edu
Abstract:
There has been much debate about the relative merits of mixed effects and population-averaged logistic models. We present a different perspective on this issue by noting that the investigation of the relationship between these models for a given dataset offers a type of sensitivity analysis that may reveal problems with assumptions of the mixed effects and/or population-averaged models for clustered binary response data in general and longitudinal binary outcomes in particular. We present several datasets in which the following violations of assumptions are associated with departures from the expected theoretical relationship between these two models: 1) negative intra-cluster correlations; 2) confounding of the response-covariate relationship by cluster effects; and 3) confounding of autoregressive relationships by the link between baseline outcomes and subject effects. Under each of these conditions, the expected theoretical attenuation of the population-averaged odds ratio relative to the cluster-specific odds ratio does not necessarily occur. In all cases, the naive fitting of a random intercept logistic model appears to lead to bias. In response, the random intercept model is modified to accommodate negative intra-cluster correlations, confounding due to clusters, or baseline correlations with random effects. Comparisons are made with GEE estimation of population-averaged models and conditional likelihood estimation of cluster-specific models. Several examples, including a cross-over trial, a multicentre nonrandomized treatment study, and a longitudinal observational study are used to illustrate these modifications.
Related Concept Videos
Standard Deviation of Calculated Results
A broad Gaussian distribution curve has a wider standard deviation, representing a data set with...
Variation: Normal Distribution, Range, and Standard Deviation
Distributions to Estimate Population Parameter
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Sampling Distribution
Standard Deviation