Related Experiment Video
Updated: May 26, 2026

10:05
Measurement & Analysis of the Temporal Discrimination Threshold Applied to Cervical Dystonia
Published on: January 27, 2018
Statistical significance of threading scores
Afshin Fayyaz Movaghar1, Guillaume Launay, Sophie Schbath
1Mathématique, Informatique, et Génome, INRA, Jouy-en-Josas, France.
Summary
We developed a method to assess protein threading score significance. By comparing scores to a Weibull distribution of random sequences, small p-values indicate significant results.
Area of Science:
- Computational Biology
- Bioinformatics
- Structural Biology
Background:
- Assessing the statistical significance of protein threading scores is crucial for accurately predicting protein structures.
- Current methods may lack a robust statistical foundation for evaluating sequence-structure alignment significance.
Purpose of the Study:
- To introduce a general and statistically rigorous method for assessing protein threading score significance.
- To establish a theoretical framework for the distribution of threading scores from random sequences.
Main Methods:
- Developed a method comparing protein sequence threading scores against a reference distribution.
- Modeled the reference distribution using a Weibull extreme value distribution based on protein contact map properties.
- Estimated distribution parameters offline using simulated sequences and interpolated for query-specific lengths.
Main Results:
- Demonstrated that threading scores of random sequences follow a Weibull extreme value distribution.
- Showcased that distribution parameters are dependent on threading method, structure, and sequence length.
- Enabled rapid computation of p-values for assessing score significance.
Conclusions:
- The proposed method provides a statistically sound approach for evaluating protein threading significance.
- The Weibull distribution offers a reliable reference for random sequence alignments.
- This facilitates more accurate protein structure prediction and analysis.
Related Concept Videos
Statistical Significance
Once data is collected from both the experimental and the control groups, a statistical analysis is conducted to find out if there are meaningful differences between the two groups. A statistical analysis determines how likely any difference found is due to chance (and thus not meaningful). In psychology, group differences are considered meaningful, or significant, if the odds that these differences occurred by chance alone are 5 percent or less. Stated another way, if we repeated this...
Comparing Experimental Results: Student's t-Test
The t-test is a statistical method used to compare the sample mean with a population mean or compare two means from two data sets. The test statistic is calculated from the standard deviation, mean, and number of measurements in the data set at a selected confidence interval and then compared to a table of critical values at this confidence level. If the test statistic is smaller than the critical value, the null hypothesis is accepted. In this case, we state that the difference between the...
Significance Testing: Overview
Significance testing is a set of statistical methods used to test whether a claim about a parameter is valid. In analytical chemistry, significance testing is used primarily to determine whether the difference between two values comes from determinate or random errors. The effect of a particular change in the measurement protocol, analyst, or sample itself can cause a deviation from the expected result. In the case of a suspected deviation/outlier, we need to be able to confirm mathematically...
Kendall's Tau Test
Kendall's tau test, also known as the Kendall rank coefficient test, is a nonparametric method for assessing association between two variables. This test is particularly useful for identifying significant correlations when the distributions of the sample and population are unknown. Developed in 1938 by the British statistician Sir Maurice George Kendall, the tau coefficient (denoted as τ) serves as a rank correlation coefficient, with values ranging from -1 to +1.
A τ value of +1 indicates that...
A τ value of +1 indicates that...
Identifying Statistically Significant Differences: The F-Test
The F-test is used to compare two sample variances to each other or compare the sample variance to the population variance. It is used to decide whether an indeterminate error can explain the difference in their values. The underlying assumptions that allow the use of the F-test include the data set or sets are normally distributed, and the data sets are independent of each other. The test statistic F is calculated by dividing one variance by another. In other words, the square of one standard...
Regression Toward the Mean
Regression toward the mean (“RTM”) is a phenomenon in which extremely high or low values—for example, and individual’s blood pressure at a particular moment—appear closer to a group’s average upon remeasuring. Although this statistical peculiarity is the result of random error and chance, it has been problematic across various medical, scientific, financial and psychological applications. In particular, RTM, if not taken into account, can interfere when researchers try to extrapolate results...

