Related Experiment Video
Updated: Feb 17, 2026

Author Spotlight: Innovations in iTUG Test for Enhanced Risk Assessment and Cognitive Insights
Published on: October 25, 2024
High-Stakes Collaborative Testing: Why Not?
Ruth E Levine1, Nicole J Borges2, Brenda J B Roman3
1a Office of Clinical Education and Department of Psychiatry and Behavioral Sciences , The University of Texas Medical Branch , Galveston , Texas , USA.
Abstract:
Phenomenon: Studies of high-stakes collaborative testing remain sparse, especially in medical education. We explored high-stakes collaborative testing in medical education, looking specifically at the experiences of students in established and newly formed teams.
Approach:
Third-year psychiatry students at 5 medical schools across 6 sites participated, with 4 participating as established team sites and 2 as comparison team sites. For the collaborative test, we used the National Board of Medical Examiners Psychiatry subject test, administering it via a 2-stage process. Students at all sites were randomly selected to participate in a focus group, with 8-10 students per site (N = 49). We also examined quantitative data for additional triangulation.
Findings:
Students described a range of heightened emotions around the collaborative test yet perceived it as valuable regardless if they were in established or newly formed teams. Students described learning about the subject matter, themselves, others, and interpersonal dynamics during collaborative testing. Triangulation of these results via quantitative data supported these themes. Insights: Despite student concerns, high-stakes collaborative tests may be both valuable and feasible. The data suggest that high-stakes tests (tests of learning or summative evaluation) could also become tests for learning or formative evaluation. The paucity of research into this methodology in medical education suggests more research is needed.
Related Concept Videos
Statistical Hypothesis Testing
Statistical significance measures the probability that an observed result occurred by chance. If this probability, known as...
Reliability and Validity
Multiple Comparison Tests
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
Significance Testing: Overview
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
Types of Hypothesis Testing
When the null and alternative hypotheses are stated, it is observed that the null hypothesis is a neutral statement against which the alternative hypothesis is tested. The alternative hypothesis is a claim that instead has a certain direction. If the null hypothesis claims that p = 0.5, the alternative hypothesis would be an opposing statement to this and can be put either p > 0.5, p < 0.5, or p...

