Necessary but Insufficient: Why Measurement Invariance Tests Need Online Probing as a Complementary Tool

Katharina Meitinger1

  • 1Katharina Meitinger is a researcher at GESIS Leibniz Institute for the Social Sciences, Mannheim, Germany, and a teaching associate at the University of Mannheim, Mannheim, Germany. The author thanks Michael Braun, Eldad Davidov, three anonymous reviewers, and the editors for their helpful comments on earlier versions of this manuscript, as well as Dorothée Behr, Lars Kaczmirek, and Wolfgang Bandilla for sharing their expertise in online probing. This research was funded by the German Research Foundation (DFG) as part of the project "Optimizing Probing Procedures for Cross-National Web Surveys" [BR 908/5-1 to Michael Braun, Wolfgang Bandilla, and Lars Kaczmirek]. An earlier version of this paper was presented at the 2016 Conference of the World Association for Public Opinion Research and won the Janet A. Harkness Student Paper Award and the NCHS Monroe Sirken Innovative Award for Young Scholars of Question Evaluation.

Related Concept Videos

Reliability and Validity01:29

Reliability and Validity

Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
14.2K
Measures of Intelligence01:29

Measures of Intelligence

Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
8.7K
Design Example: Measuring Distance Between Two Points with Obstructions01:10

Design Example: Measuring Distance Between Two Points with Obstructions

When measuring distances in areas with physical obstructions, such as a lake in a field, surveyors must employ techniques to calculate accurate lengths without direct line measurements. One effective method is the offset technique, which allows for precise distance estimation over inaccessible stretches.In this scenario, a surveyor must measure a side of an area that crosses a lake. Since the measuring tape cannot span the lake, the surveyor begins by establishing a baseline that aligns with...
452
Testing a Claim about Population Proportion01:24

Testing a Claim about Population Proportion

A complete procedure for testing a claim about a population proportion is provided here.
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
4.0K
Statistical Analysis: Overview01:11

Statistical Analysis: Overview

When we take repeated measurements on the same or replicated samples, we will observe inconsistencies in the magnitude. These inconsistencies are called errors. To categorize and characterize these results and their errors, the researcher can use statistical analysis to determine the quality of the measurements and/or suitability of the methods.
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
16.7K
Multiple Comparison Tests01:13

Multiple Comparison Tests

Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
4.5K