OSVE或多项选择测试:这是一个相关的问题吗?
Francine Jomara Lopes1, Renato Fraga Righetti1, Matheus Belloni Torsani2
1Hospital Sírio-Libanês, São Paulo, São Paulo, SP, Brazil; Department of Clinical Medicine, Faculdade de Medicina, Universidade de São Paulo (FMUSP), São Paulo, SP, Brazil.
Clinics (Sao Paulo, Brazil)
|January 26, 2025
概括
客观结构化临床检查 (OSCE) 的替代方案,客观结构化虚拟检查 (OSVE),与多项选择测试 (MCT) 和平均成绩点 (GPA) 显示轻度至中度的相关性. 这些发现表明,OSVE和MCT可以在医学教育中的形成性评估中互换.
科学领域:
- 医学教育 医学教育
- 临床技能评估 临床技能评估
- 虚拟评估工具 虚拟评估工具
背景情况:
- 传统的客观结构化临床检查 (OSCE) 对于评估临床技能至关重要,但在COVID-19大流行期间面临挑战.
- 目标结构虚拟考试 (OSVE) 作为一种替代的评估方法出现了.
- 本研究评估了OSVE与其他学术指标的相关性.
研究的目的:
- 为了将目标结构虚拟考试 (OSVE) 的分数与多选项测试 (MCT) 的结果相关联.
- 评估OSVE分数和职务人员平均成绩 (GPA) 之间的关系.
- 确定OSVE作为医学培训中的形成性评估工具的适用性.
主要方法:
- 一项涉及129个职场的横截面研究.
- 两个OSVE和两个MCT的比较,涵盖第5年和第6年课程内容.
- 与最终毕业成绩 (GPA) 相对应的相关性分析.
主要成果:
- 在OSVE和GPA之间发现了正相关性 (第5年R=0.481,第6年R=0.439).
- 在MCT-5th和GPA (R=0.681) 之间观察到中等相关性.
- 在OSVE和MCT得分之间注意到轻度至中度的相关性,但有一些例外 (例如,OSVE-6和MCT-6).
结论:
- OSVE,MCT和GPA之间的相关性表明它们可以互换地用于形成性评估.
- OSVE 和 MCT 是可行的和有效的工具来加强内部培训和监测.
- 虚拟评估方法为医学教育的连续性提供了一个可行的替代方案.
相关概念视频
Multiple Comparison Tests
3.8K
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
3.8K
Reliability and Validity
12.7K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
12.7K
Significance Testing: Overview
3.3K
Significance testing is a set of statistical methods used to test whether a claim about a parameter is valid. In analytical chemistry, significance testing is used primarily to determine whether the difference between two values comes from determinate or random errors. The effect of a particular change in the measurement protocol, analyst, or sample itself can cause a deviation from the expected result. In the case of a suspected deviation/outlier, we need to be able to confirm mathematically...
3.3K
Multiple Regression
2.9K
Multiple regression assesses a linear relationship between one response or dependent variable and two or more independent variables. It has many practical applications.
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
2.9K
Goodness-of-Fit Test
3.3K
The goodness-of-fit test is a type of hypothesis test which determines whether the data "fits" a particular distribution. For example, one may suspect that some anonymous data may fit a binomial distribution. A chi-square test (meaning the distribution for the hypothesis test is chi-square) can be used to determine if there is a fit. The null and alternative hypotheses may be written in sentences or stated as equations or inequalities. The test statistic for a goodness-of-fit test is given as...
3.3K
Detection of Gross Error: The Q Test
5.6K
When one or more data points appear far from the rest of the data, there is a need to determine whether they are outliers and whether they should be eliminated from the data set to ensure an accurate representation of the measured value. In many cases, outliers arise from gross errors (or human errors) and do not accurately reflect the underlying phenomenon. In some cases, however, these apparent outliers reflect true phenomenological differences. In these cases, we can use statistical methods...
5.6K


