Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Multiple Comparison Tests01:13

Multiple Comparison Tests

4.5K
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
4.5K
Behrens–Fisher Test00:57

Behrens–Fisher Test

309
The Behrens-Fisher test is a statistical method designed to address the Behrens-Fisher problem, which arises when comparing the means of two normally distributed populations with unequal variances. Unlike the Student's t-test, which assumes equal variances, the Behrens-Fisher test allows for mean comparison without this restrictive assumption. This flexibility makes it particularly valuable in scenarios where two independent samples exhibit normality but lack variance homogeneity.
This test...
309
Bonferroni Test01:10

Bonferroni Test

3.5K
The Bonferroni test is a statistical test named after Carlo Emilio Bonferroni, an Italian mathematician best known for Bonferroni inequalities. This statistical test is a type of multiple comparison test to determine which means are different than the rest. Bonferroni test can minimize the Type 1 error by reducing the significance level alpha, which otherwise increases with sample pairs.
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
3.5K
Friedman Two-way Analysis of Variance by Ranks01:21

Friedman Two-way Analysis of Variance by Ranks

530
Friedman's Two-Way Analysis of Variance by Ranks is a nonparametric test designed to identify differences across multiple test attempts when traditional assumptions of normality and equal variances do not apply. Unlike conventional ANOVA, which requires normally distributed data with equal variances, Friedman's test is ideal for ordinal or non-normally distributed data, making it particularly useful for analyzing dependent samples, such as matched subjects over time or repeated measures...
530
One-Way ANOVA: Equal Sample Sizes01:15

One-Way ANOVA: Equal Sample Sizes

4.3K
One-Way ANOVA can be performed on three or more samples with equal or unequal sample sizes. When one-way ANOVA is performed on two datasets with samples of equal sizes, it can be easily observed that the computed F statistic is highly sensitive to the sample mean.
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
4.3K
Self-Report Tests of Personality01:22

Self-Report Tests of Personality

1.1K
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
1.1K

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

Detection of high-grade dysplasia and adenocarcinoma in Barrett's esophagus using high-resolution virtual chromoendoscopy versus the Seattle protocol: the CONVERSE study.

BMC gastroenterology·2026
Same author

Development and psychometric validation of the PERQOLATEUR questionnaire: an individualized self-report measure of the impact of chronic illness and treatment in clinical practice.

Journal of patient-reported outcomes·2026
Same author

Association between difficulties navigating the French healthcare system and healthcare utilisation: results from the National Health Literacy Survey (HLS<sub>19</sub>).

BMJ open·2026
Same author

Trends in prevalence and disability burden of main rheumatic and musculoskeletal diseases in France 2008-2022.

RMD open·2026
Same author

From coming to terms with the gambling problems to post-traumatic growth: A novel nine-stage theoretical model of recovery from gambling disorder.

Journal of behavioral addictions·2026
Same author

Sex Differences in Social, Health, and Lifestyle Characteristics Associated With Binge-Eating Behaviors: Results From a French National Random Population-Based Study.

The International journal of eating disorders·2026

Related Experiment Video

Updated: Mar 9, 2026

Applying an eMASS Customization Program as a Research Tool to Evaluate Consumer Benefits
08:27

Applying an eMASS Customization Program as a Research Tool to Evaluate Consumer Benefits

Published on: September 27, 2019

7.3K

Differential Item Functioning (DIF) and Subsequent Bias in Group Comparisons using a Composite Measurement Scale: A

Alexandra Rouquette1, Jean-Benoit Hardouin, Joel Coste

  • 1Alexandra Rouquette, Hotel-Dieu Hospital, Biostatistics and Epidemiology Department, 1 place du parvis Notre-Dame, 75181 Paris cedex 04, France, alex.rouquette@gmail.com.

Journal of Applied Measurement
|December 28, 2016
PubMed
Summary

Differential Item Functioning (DIF) can bias group comparisons in composite scales. Bias is mainly influenced by DIF size and proportion for uniform DIF, while non-uniform DIF effects are minimal.

More Related Videos

Multimedia Battery for Assessment of Cognitive and Basic Skills in Mathematics BM-PROMA
10:58

Multimedia Battery for Assessment of Cognitive and Basic Skills in Mathematics BM-PROMA

Published on: August 28, 2021

5.1K
Author Spotlight: Validation of SICOLE-R for Assessing Cognitive and Reading Skills in Spanish-Speaking Children and Its Role in Personalized Education
09:00

Author Spotlight: Validation of SICOLE-R for Assessing Cognitive and Reading Skills in Spanish-Speaking Children and Its Role in Personalized Education

Published on: August 16, 2024

1.3K

Related Experiment Videos

Last Updated: Mar 9, 2026

Applying an eMASS Customization Program as a Research Tool to Evaluate Consumer Benefits
08:27

Applying an eMASS Customization Program as a Research Tool to Evaluate Consumer Benefits

Published on: September 27, 2019

7.3K
Multimedia Battery for Assessment of Cognitive and Basic Skills in Mathematics BM-PROMA
10:58

Multimedia Battery for Assessment of Cognitive and Basic Skills in Mathematics BM-PROMA

Published on: August 28, 2021

5.1K
Author Spotlight: Validation of SICOLE-R for Assessing Cognitive and Reading Skills in Spanish-Speaking Children and Its Role in Personalized Education
09:00

Author Spotlight: Validation of SICOLE-R for Assessing Cognitive and Reading Skills in Spanish-Speaking Children and Its Role in Personalized Education

Published on: August 16, 2024

1.3K

Area of Science:

  • Psychometrics
  • Statistical Modeling

Background:

  • Composite measurement scales are widely used to assess complex constructs.
  • Differential Item Functioning (DIF) can introduce bias in group comparisons when not accounted for.

Purpose of the Study:

  • To investigate the conditions under which estimating group differences using composite scales is biased due to undetected DIF.
  • To quantify measurement bias in realistic scenarios of composite scale application.

Main Methods:

  • Simulated 642 datasets using the Partial Credit Model.
  • Employed Analysis of Variance (ANOVA) to assess the impact of seven factors on bias.
  • Factors included sample size, true group difference, scale length, DIF proportion, DIF size, item location, and DIF type (uniform/non-uniform).

Main Results:

  • For uniform DIF, DIF size and the proportion of items exhibiting DIF, along with their interaction, significantly impacted bias.
  • The influence of non-uniform DIF on bias was found to be negligible.

Conclusions:

  • Measurement bias due to DIF in composite scales is quantifiable under various realistic conditions.
  • Understanding the impact of uniform DIF is crucial for accurate group comparisons.