Related Experiment Video
Updated: Aug 5, 2026

13:55
Combined Immunofluorescence and DNA FISH on 3D-preserved Interphase Nuclei to Study Changes in 3D Nuclear Organization
Published on: February 3, 2013
Non-independence in statistical tests for discrete cross-species data
Journal of Theoretical Biology
|February 21, 1998
Summary
This study reveals three biases in statistical tests for character state associations in cross-species data. The "family problem" affects all methods, interacting with reconstruction biases to impact phylogenetic analyses.
Area of Science:
- Phylogenetics
- Statistical methods
- Evolutionary biology
Background:
- Statistical tests for character state associations in cross-species data are crucial for evolutionary inference.
- Existing methods may be susceptible to biases arising from data structure and reconstruction techniques.
- Understanding these biases is essential for accurate phylogenetic analyses.
Purpose of the Study:
- To identify and describe previously undetected biases in statistical tests for character state associations.
- To investigate the 'family problem' and its interaction with non-independence in character state reconstruction.
- To evaluate the impact of these biases on phylogenetic data analysis.
Main Methods:
- The study identifies biases arising from non-independence in statistical tests.
- It introduces the 'family problem,' a general bias affecting character state reconstruction.
- It analyzes biases specific to joint and single character state reconstruction methods.
Main Results:
- Three previously undetected biases in statistical tests for character state associations were identified.
- The 'family problem' is a general bias affecting all tested methods.
- Biases in character reconstruction methods interact with the family problem, influencing results.
Conclusions:
- The family problem significantly impacts the reliability of statistical tests for character state associations.
- The choice of character reconstruction method influences the type and magnitude of bias observed.
- Accurate phylogenetic inference requires careful consideration of these identified biases and their interactions.
Related Concept Videos
Introduction to Test of Independence
In statistics, the term independence means that one can directly obtain the probability of any event involving both variables by multiplying their individual probabilities. Tests of independence are chi-square tests involving the use of a contingency table of observed (data) values.
The test statistic for a test of independence is similar to that of a goodness-of-fit test:
The test statistic for a test of independence is similar to that of a goodness-of-fit test:
Hypothesis Test for Test of Independence
The test of independence is a chi-square-based test used to determine whether two variables or factors are independent or dependent. This hypothesis test is used to examine the independence of the variables. One can construct two qualitative survey questions or experiments based on the variables in a contingency table. The goal is to see if the two variables are unrelated (independent) or related (dependent). The null and alternative hypotheses for this test are:
H0: The two variables (factors)...
H0: The two variables (factors)...
Test for Homogeneity
The goodness–of–fit test can be used to decide whether a population fits a given distribution, but it will not suffice to decide whether two populations follow the same unknown distribution. A different test, called the test for homogeneity, can be used to conclude whether two populations have the same distribution. To calculate the test statistic for a test for homogeneity, follow the same procedure as with the test of independence. The hypotheses for the test for homogeneity can be stated as...
One-Way ANOVA: Equal Sample Sizes
One-Way ANOVA can be performed on three or more samples with equal or unequal sample sizes. When one-way ANOVA is performed on two datasets with samples of equal sizes, it can be easily observed that the computed F statistic is highly sensitive to the sample mean.
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Kruskal-Wallis Test
The Kruskal-Wallis test, also known as the Kruskal-Wallis H test, serves as a nonparametric alternative to the one-way ANOVA, offering a solution for analyzing the differences across three or more independent groups based on a single, ordinal-dependent variable. This statistical test is particularly valuable in scenarios where the data does not meet the normal distribution assumption required by its parametric counterparts. Kruskal-Wallis test is designed typically to handle ordinal data or...
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Statistical inference techniques, paramount in hypothesis testing, differentiate into two broad categories: parametric and nonparametric statistics.
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance, comparing...
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance, comparing...

