Related Experiment Video
Updated: Mar 10, 2026

07:35
Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
8.1K
Inferences about competing measures based on patterns of binary significance tests are questionable
Patrick E Shrout1, Marika Yip-Bannicq1
1Department of Psychology, New York University.
Psychological Methods
|December 13, 2016
Summary
Binary significance tests (BST) for incremental validity can be misleading, leading to false conclusions up to 30% of the time. This study demonstrates flawed reasoning in validity studies and offers better methods for accurate inference.
Area of Science:
- Psychological measurement
- Statistical inference
- Research methodology
Background:
- Demonstrating the validity of new measures is crucial, often involving incremental validity testing against existing measures.
- Regression methods and binary significance tests (BST) are commonly used to argue for a new measure's incremental validity.
- The traditional BST approach relies on specific patterns of statistical significance when measures are jointly tested.
Purpose of the Study:
- To evaluate the reliability of arguments for incremental validity based on binary significance tests (BST).
- To identify the conditions under which BST can lead to erroneous conclusions in validity studies.
- To propose and illustrate alternative, more robust methods for assessing measure validity and incremental contribution.
Main Methods:
- Analysis of statistical inference in regression models, specifically focusing on binary significance testing (BST).
- Simulation or examination of scenarios with modest statistical power to assess the rate of false conclusions from BST.
- Application of alternative inferential methods to real-world data, using construal level as a case study.
Main Results:
- Binary significance test (BST) arguments for incremental validity can result in incorrect conclusions in up to 30% of cases, particularly with modest statistical power.
- The 'black and white' approach of relying solely on significance/non-significance is shown to be misleading.
- Alternative methods provide stronger and more accurate inferences about measure validity and added information.
Conclusions:
- Relying on binary significance tests for incremental validity can lead to substantial error rates.
- Researchers should adopt more nuanced statistical approaches for evaluating new measures.
- Accurate assessment of measure validity and incremental contribution is essential for advancing psychological science.
More Related Videos
Related Concept Videos
Significance Testing: Overview
12.9K
Significance testing is a set of statistical methods used to test whether a claim about a parameter is valid. In analytical chemistry, significance testing is used primarily to determine whether the difference between two values comes from determinate or random errors. The effect of a particular change in the measurement protocol, analyst, or sample itself can cause a deviation from the expected result. In the case of a suspected deviation/outlier, we need to be able to confirm mathematically...
12.9K
Bonferroni Test
3.5K
The Bonferroni test is a statistical test named after Carlo Emilio Bonferroni, an Italian mathematician best known for Bonferroni inequalities. This statistical test is a type of multiple comparison test to determine which means are different than the rest. Bonferroni test can minimize the Type 1 error by reducing the significance level alpha, which otherwise increases with sample pairs.
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
3.5K
Introduction to the Sign Test
1.4K
The sign test is an important tool in nonparametric statistics, offering a straightforward yet effective method for analyzing matched pairs, nominal data, or hypotheses concerning the median of a population. It transforms data points into positive or negative signs, avoiding the need for assumptions about data distribution and instead focusing on the direction of change. It is particularly valuable when data does not conform to the normal distribution requirements of many parametric tests. For...
1.4K
Sign Test for Matched Pairs
448
The sign test for matched pairs offers a robust method for comparing two paired samples, often for the effects of an intervention in one of them. This method is very useful in situations where the underlying distribution of the data is unknown. The test compares two related samples—often pre- and post-treatment measurements on the same subjects—to determine if there are significant differences in their median values.
To conduct the sign test, we first calculate the differences in...
To conduct the sign test, we first calculate the differences in...
448
Sign Test for Nominal Data
426
The sign test is a nonparametric method used to evaluate hypotheses about the median of a single sample or to compare the medians of two related samples. The sign test is particularly useful when dealing with nominal data, which includes distinct categories without an inherent order, such as names, labels, and preferences. Nominal data restricts statistical analysis to evaluating population proportions rather than mean or median values that require continuous data.
For example, consider a...
For example, consider a...
426
Comparing Experimental Results: Student's t-Test
6.2K
The t-test is a statistical method used to compare the sample mean with a population mean or compare two means from two data sets. The test statistic is calculated from the standard deviation, mean, and number of measurements in the data set at a selected confidence interval and then compared to a table of critical values at this confidence level. If the test statistic is smaller than the critical value, the null hypothesis is accepted. In this case, we state that the difference between the...
6.2K

