Related Experiment Video
Updated: Oct 1, 2025

Development of an Individual-Tree Basal Area Increment Model using a Linear Mixed-Effects Approach
Published on: July 3, 2020
Semiparametric marginal methods for clustered data adjusting for informative cluster size with nonignorable zeros
Biyi Shen1, Chixiang Chen2, Vernon M Chinchilli1
1Division of Biostatistics and Bioinformatics, Department of Public Health Sciences, Penn State College of Medicine, Hershey, PA, USA.
Abstract:
Clustered or longitudinal data are commonly encountered in clinical trials and observational studies. This type of data could be collected through a real-time monitoring scheme associated with some specific event, such as disease recurrence, hospitalization, or emergency room visit. In these contexts, the cluster size could be informative because of its potential correlation with disease status, since more frequency of observations may indicate a worsening health condition. However, for some clusters/subjects, there are no measures or relevant medical records. Under such circumstances, these clusters/subjects may have a considerably lower risk of an event occurrence or may not be susceptible to such events at all, indicating a nonignorable zero cluster size. There is a substantial body of literature using observations from those clusters with a nonzero informative cluster size only, but few works discuss informative nonignorable zero-sized clusters. To utilize the information from both event-free and event-occurring participants, we propose a weighted within-cluster-resampling (WWCR) method and its asymptotically equivalent method, dual-weighted generalized estimating equations (WWGEE) by adopting the inverse probability weighting technique. The asymptotic properties are rigorously presented theoretically. Extensive simulations and an illustrative example of the Assessment, Serial Evaluation, and Subsequent Sequelae of Acute Kidney Injury (ASSESS-AKI) study are performed to analyze the finite-sample behavior of our methods and to show their advantageous performance compared to the existing approaches.
More Related Videos
12:27Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations
Published on: February 15, 2017
08:27Applying an eMASS Customization Program as a Research Tool to Evaluate Consumer Benefits
Published on: September 27, 2019
Related Concept Videos
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Parametric Survival Analysis: Weibull and Exponential Methods
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
Estimating Population Mean with Unknown Standard Deviation
William S. Gosset (1876–1937) of the...
Sampling Plans
Random sampling is a method where each member of the population has an equal chance of being selected for the sample. It involves selecting individuals randomly, often using random number generators or lottery-type methods. For example, when analyzing the properties of a...
One-Way ANOVA: Equal Sample Sizes
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...