对集群随机测试负面设计的随机化干扰与对登革研究的应用:无偏见的估计,部分遵守和阶段性设计
Bingkai Wang1, Suzanne M Dufault2, Dylan S Small1
1Department of Statistics and Data Science, The Wharton School, University of Pennsylvania.
The annals of applied statistics
|August 26, 2025
概括
集群随机测试负面设计 (CR-TND) 提供了成本高效的登革热监测. 提出了新的方法来纠正CR-TND在寻求医疗行为的变化时的偏差,从而改善登革热控制策略.
科学领域:
- 流行病学和生物统计学
- 传染病控制
背景情况:
- 登革热是世界卫生组织确定的一大全球性健康威胁.
- 应用沃尔巴基亚消除登革热 (AWED) 研究使用了一种新的集群随机测试负面设计 (CR-TND) 来控制登革热.
- 与传统的集群随机试验相比,CR-TND通过被动监测提供了成本效益.
研究的目的:
- 在一个强大的随机推断框架下调查CR-TND的统计假设和属性.
- 在CR-TND分析中解决潜在的偏差和膨胀的I型错误,当医疗寻求行为在集群中不同时.
- 提出和验证一种用于CR-TND准确推断的新统计方法.
主要方法:
- 使用随机推断框架分析CR-TND.
- 开发和应用日志对比估计器以调整共变量和不同的医疗寻求行为.
- 扩展方法以适应部分干预合规和阶梯设计.
- 通过模拟研究和重新分析AWED研究数据进行验证.
主要成果:
- 在当前CR-TND分析方法中发现偏差和膨胀的I型错误,当不同的医疗寻求行为在集群之间有所不同时.
- 拟议的日志对比估计器有效消除了偏差并提高了精度.
- 扩展方法成功处理部分合规和阶梯设计.
结论:
- 拟议的日志对比估计器提高了登革热监测CR-TND的有效性和可靠性.
- 开发的统计框架为复杂的集群随机试验设计提供了强大的推断.
- 这些进展有助于控制登革热和类似传染病的更有效和更经济的策略.
相关概念视频
Randomized Experiments
7.2K
The randomization process involves assigning study participants randomly to experimental or control groups based on their probability of being equally assigned. Randomization is meant to eliminate selection bias and balance known and unknown confounding factors so that the control group is similar to the treatment group as much as possible. A computer program and a random number generator can be used to assign participants to groups in a way that minimizes bias.
Simple randomization
Simple...
Simple randomization
Simple...
7.2K
Group Design
9.6K
The most basic experimental design involves two groups: the experimental group and the control group. The two groups are designed to be the same except for one difference— experimental manipulation. The experimental group gets the experimental manipulation—that is, the treatment or variable being tested—and the control group does not. Since experimental manipulation is the only difference between the experimental and control groups, we can be sure that any differences between...
9.6K
McNemar's Test
415
McNemar's Test is a nonparametric statistical test used to determine if there is a significant difference in proportions between two related groups when the outcome is binary (e.g., yes/no, success/failure). It is beneficial when we have paired data, such as pre-test/post-test designs, where the same subjects are measured under two different conditions. The test is named after the statistician Quinn McNemar, who introduced it in 1947. It is commonly used in situations where subjects are...
415
Wald-Wolfowitz Runs Test II
316
The Wald-Wolfowitz runs test, commonly referred to as the runs test, is a nonparametric test used to assess the randomness of ordered data. The test evaluates the number of runs, which are consecutive sequences of similar elements within the data. If the number of runs is significantly higher or lower than expected, the data is considered non-random, indicating a detectable pattern or structure.
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and...
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and...
316
Introduction to Test of Independence
2.4K
In statistics, the term independence means that one can directly obtain the probability of any event involving both variables by multiplying their individual probabilities. Tests of independence are chi-square tests involving the use of a contingency table of observed (data) values.
The test statistic for a test of independence is similar to that of a goodness-of-fit test:
The test statistic for a test of independence is similar to that of a goodness-of-fit test:
2.4K
The Anderson-Darling Test
873
The Anderson-Darling test is a statistical method used to determine whether a data sample is likely drawn from a specific theoretical distribution. Unlike parametric tests, it does not require assumptions about specific parameters of the distribution. Instead, it compares the sample's empirical cumulative distribution function (ECDF) with the cumulative distribution function (CDF) of the hypothesized distribution. Critical values for the test are specific to the chosen distribution rather...
873


