Optimal experimental designs for estimating genetic and non-genetic effects underlying infectious disease
Christopher Pooley1,2, Glenn Marion3, Stephen Bishop
1Biomathematics and Statistics Scotland, James Clerk Maxwell Building, The King's Buildings, Peter Guthrie Tait Road, Edinburgh, EH9 3FD, UK. chris.pooley@bioss.ac.uk.
Background:
The spread of infectious diseases in populations is controlled by the susceptibility (propensity to acquire infection), infectivity (propensity to transmit infection), and recoverability (propensity to recover/die) of individuals. Estimating genetic risk factors for these three underlying host epidemiological traits can help reduce disease spread through genetic control strategies. Previous studies have identified important 'disease resistance single nucleotide polymorphisms (SNPs)', but how these affect the underlying traits is an unresolved question. Recent advances in computational statistics make it now possible to estimate the effects of SNPs on host traits from epidemic data (e.g. infection and/or recovery times of individuals or diagnostic test results). However, little is known about how to effectively design disease transmission experiments or field studies to maximise the precision with which these effects can be estimated.
Results:
In this paper, we develop and validate analytical expressions for the precision of the estimates of SNP effects on the three above host traits for a disease transmission experiment with one or more non-interacting contact groups. Maximising these expressions leads to three distinct 'experimental' designs, each specifying a different set of ideal SNP genotype compositions across groups: (a) appropriate for a single contact-group, (b) a multi-group design termed "pure", and (c) a multi-group design termed "mixed", where 'pure' and 'mixed' refer to groupings that consist of individuals with uniformly the same or different SNP genotypes, respectively. Precision estimates for susceptibility and recoverability were found to be less sensitive to the experimental design than estimates for infectivity. Whereas the analytical expressions suggest that the multi-group pure and mixed designs estimate SNP effects with similar precision, the mixed design is preferred because it uses information from naturally-occurring rather than artificial infections. The same design principles apply to estimates of the epidemiological impact of other categorical fixed effects, such as breed, line, family, sex, or vaccination status. Estimation of SNP effect precisions from a given experimental setup is implemented in an online software tool SIRE-PC.
Conclusions:
Methodology was developed to aid the design of disease transmission experiments for estimating the effect of individual SNPs and other categorical variables that underlie host susceptibility, infectivity and recoverability. Designs that maximize the precision of estimates were derived.
More Related Videos
Related Concept Videos
Study Designs in Epidemiology
Observational studies are those where the researcher does not intervene but rather observes natural variations. They include cross-sectional, cohort, and...
What is an Experiment?
Introduction to Epidemiology
Causality in Epidemiology
Statistical Methods for Analyzing Epidemiological Data
Study Design in Statistics
Does aspirin reduce the risk of heart attacks? Is one brand of fertilizer more effective at growing roses than another? Is fatigue as dangerous to a driver as the influence of alcohol? Questions like these are answered using randomized experiments with proper...


