Related Experiment Video
Updated: Aug 28, 2025

08:33
A Cross-Disciplinary and Multi-Modal Experimental Design for Studying Near-Real-Time Authentic Examination Experiences
Published on: September 4, 2019
7.1K
Multiple, speeded assessments under scrutiny: Underlying theory, design considerations, reliability, and validity
Christoph N Herde1, Filip Lievens1
1Lee Kong Chian School of Business, Singapore Management University.
The Journal of Applied Psychology
|September 15, 2022
Summary
Multiple speeded assessments can reliably predict future job performance, but only when aggregated across many situations and raters. Single, short assessments are not recommended for employee selection.
Area of Science:
- Organizational Psychology
- Human Resources Management
- Assessment Psychology
Background:
- Speeded assessments are increasingly used in employee selection.
- There is a lack of conceptual and empirical evidence supporting their reliability and validity.
- This study addresses the effectiveness of these rapid assessment methods.
Purpose of the Study:
- To conceptualize multiple speeded behavioral job simulations.
- To investigate the reliability and validity of these assessments for predicting future performance.
- To examine the impact of rating aggregation and assessor training on validity.
Main Methods:
- Conceptualization using stimulus and response domain sampling.
- Application of the thin slices of behavior paradigm.
- Two studies involving 96 MBA students assessed in 18 speeded role-plays (3-min each).
- Analysis of aggregated ratings across multiple assessors and role-plays.
Main Results:
- Individual speeded role-plays lacked reliability and validity.
- Aggregated scores across all role-plays and assessors showed high predictive validity (.54).
- Validity remained high even with shorter assessment durations (1 min) and basic assessor training.
Conclusions:
- Aggregating ratings from multiple, diverse situations is crucial for comprehensive domain coverage.
- This aggregation captures both ability and personality traits (e.g., extraversion, agreeableness).
- The findings underscore the importance of domain sampling and caution against using single speeded assessments.
Related Concept Videos
Reliability and Validity
12.9K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
12.9K
Multiple Comparison Tests
4.0K
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
4.0K
Measures of Intelligence
7.8K
Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
7.8K
Factorial Design
13.1K
Factorial Analysis is an experimental design that applies Analysis of Variance (ANOVA) statistical procedures to examine a change in a dependent variable due to more than one independent variable, also known as factors. Changes in worker productivity can be reasoned, for example, to be influenced by salary and other conditions, such as skill level. One way to test this hypothesis is by categorizing salary into three levels (low, moderate, and high) and skills sets into two levels (entry level...
13.1K
Self-Report Tests of Personality
438
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
438
Data Validation
5.2K
Data validation is an essential part of a comprehensive assessment. Validation is confirming or verifying and opening the door to gathering more assessment data as it clarifies vague or unclear data. The process of checking and verifying the collected information is called data validation. The primary purpose of data validation is to ensure data is as free from error, bias, and misinterpretation as possible.
Nursing assessment guides are generally based on holistic models rather than medical...
Nursing assessment guides are generally based on holistic models rather than medical...
5.2K

