在非Rasch IRT中基于总结分数的评分中增加不确定性
1Independent Researcher, Chicago, IL, USA.
Applied psychological measurement
|June 16, 2025
概括
在患者报告结果测量信息系统 (PROMIS) 中总结得分 (SS) 评分提供了方便性,但增加了不确定性. 量化这种增加对于确定SS评分是否与响应模式 (RP) 评分一样准确至关重要.
科学领域:
- 心理测量 心理测量 心理测量
- 健康 结果 研究 研究 结果
- 项目响应理论 (IRT)
背景情况:
- 总结得分 (SS) 评分在非冲击项目响应理论 (IRT) 中被用于管理方便,特别是在像患者报告结果测量信息系统 (PROMIS) 这样的系统中.
- 目前,PROMIS使用基于SS和基于响应模式 (RP) 的评分方法.
- 与RP得分相比,SS得分引入了额外的不确定性,称为"增加".
研究的目的:
- 在患者报告结果测量信息系统 (PROMIS) 简单表格中量化与总结得分 (SS) 评分相关的不确定性增加.
- 评估这种量化增加对SS评分准确性和可行性的影响,与响应模式 (RP) 评分相比.
- 为PROMIS简单表格提供一种报告增幅的方法,以告知得分决策.
主要方法:
- 利用贝叶斯分析来确定SS和RP得分之间的关系.
- 使用差异分解来将RP得分的不确定性与SS得分的增加分开.
- 量化了两个特定的PROMIS缩短形式 (SF) 的增长.
主要成果:
- 在分析的两种PROMIS简体表格之间,量化增加的差异很大.
- 一个简短的表格显示了可以忽略的增加,这表明SS得分与RP得分一样准确.
- 另一种简短的形式显示大幅增加,质疑SS评分作为次要选项的可行性.
结论:
- 与SS评分相关的不确定性增加是影响其与RP评分相对准确性的关键因素.
- 增幅的幅度在PROMIS简体形式上有所不同,需要个别量化.
- 在实施SS评分时,报告每个简单表格的量化增加至关重要.
相关概念视频
Reliability and Validity
13.2K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
13.2K
Uncertainty in Measurement: Accuracy and Precision
82.9K
Scientists typically make repeated measurements of a quantity to ensure the quality of their findings and to evaluate both the precision and the accuracy of their results. Measurements are said to be precise if they yield very similar results when repeated in the same manner. A measurement is considered accurate if it yields a result that is very close to the true or the accepted value. Precise values agree with each other; accurate values agree with a true value.
82.9K
Propagation of Uncertainty from Random Error
1.1K
An experiment often consists of more than a single step. In this case, measurements at each step give rise to uncertainty. Because the measurements occur in successive steps, the uncertainty in one step necessarily contributes to that in the subsequent step. As we perform statistical analysis on these types of experiments, we must learn to account for the propagation of uncertainty from one step to the next. The propagation of uncertainty depends on the type of arithmetic operation performed on...
1.1K
Uncertainty: Overview
999
In analytical chemistry, we often perform repetitive measurements to detect and minimize inaccuracies caused by both determinate and indeterminate errors. Despite the cares we take, the presence of random errors means that repeated measurements almost never have exactly the same magnitude. The collective difference between these measurements - observed values - and the estimated or expected value is called uncertainty. Uncertainty is conventionally written after the estimated or expected value.
999
Introduction to z Scores
687
A z score (or standardized value) is measured in units of the standard deviation. It indicates how many standard deviations the value x is above (to the right of) or below (to the left of) the mean, μ. Values of x that are larger than the mean have positive z scores, and values of x that are smaller than the mean have negative z scores. If x equals the mean, then x has a zero z score. It is important to note that the mean of the z scores is zero, and the standard deviation is one.
z scores...
z scores...
687
Wilcoxon Rank-Sum Test
355
The Wilcoxon rank-sum test, also known as the Mann-Whitney U test, is a nonparametric test used to determine if there is a significant difference between the distributions of two independent samples. This test is designed specifically for two independent populations and has the following key requirements:
355


