相关实验视频
Updated: Jan 18, 2026

Assessment and Communication for People with Disorders of Consciousness
Published on: August 1, 2017
挑战规范:考试的长度取决于分类准确性或可靠性
Stefan K Schauber1,2, Matt Homer3
1Section for Health Sciences Education (HELP), Faculty of Medicine, University of Oslo, Norway.
分类准确性,而不是可靠性,对于医学教育考试来说更好. 这项研究表明,分类准确性建议缩短测试长度,改善可辩护的通过-失败决策,减少评估负担.
科学领域:
- 医学教育 医学教育
- 心理测量 心理测量 心理测量
- 评估科学 评估科学
背景情况:
- 传统的医学教育考试通常依赖于可靠性指数来确定测试长度.
- 然而,这些考试的主要目标是确保可辩护的通过-失败决策,这一目的仅靠可靠性无法完全实现.
研究的目的:
- 挑战在医学教育中使用可靠性指数来决定测试长度的使用.
- 建议将分类准确性作为一个更合适的度量,用于通过或失败的决定.
- 实证地证明分类准确性导致与可靠性相比,建议的测试长度更短.
主要方法:
- 从本科医学知识考试中重新采样的测试数据的分析 (N=52,500个数据集).
- 在生成的合成考试中,切割分数和测试长度的系统变化.
- 估计每个数据集的可靠性和分类准确性指数.
主要成果:
- 分类的准确性,与可靠性不同,随着通过-失败决策的分数而变化.
- 可靠性和分类准确性与测试长度有不同的关系.
- 测试可靠性的最佳测试长度是~100项,无论通过率如何.
- 对于分类准确度,50个项目可以达到95%的准确度,故障率≤5%.
结论:
- 倡导在医学教育评估中采用分类准确性,以补充现有的可靠性措施.
- 实施分类准确性可以减少候选人和开发人员的评估负担.
- 强调在通过/失败分类中考虑错误的阳性和错误的负性决定的重要性.
更多相关视频
09:00Author Spotlight: Validation of SICOLE-R for Assessing Cognitive and Reading Skills in Spanish-Speaking Children and Its Role in Personalized Education
Published on: August 16, 2024
09:18Author Spotlight: Assessing the Reliability of Doppler Ultrasound in Measuring Leg Blood Flow
Published on: December 15, 2023
相关概念视频
Receiver Operating Characteristic Plot
Accuracy and Precision
Reliability and Validity
Sensitivity, Specificity, and Predicted Value
Sensitivity is the...
Uncertainty in Measurement: Accuracy and Precision
Statistical Analysis: Overview
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...