从四合一到卡帕:如何评估二进制尺度上的可靠性
1Department of Methodology and Statistics, Care and Public Health Research Institute (CAPHRI), Maastricht University, Maastricht, The Netherlands.
The British journal of mathematical and statistical psychology
|December 8, 2025
概括
估计二进制数据的可靠性是一项挑战. 本研究比较了三种方法,发现正常近似和频率主义方法对二进制尺度可靠性不可靠. 潜变量方法提供了更好的洞察力.
科学领域:
- 心理测量 心理测量 心理测量
- 统计方法 统计方法
背景情况:
- 心理测量的可靠性对于测量的准确性至关重要.
- 与定量尺度相比,二元结果对可靠性估计提出了独特的挑战.
- 对于定量数据,建立了经典测试理论和类内相关系数.
研究的目的:
- 审查和链接三个主要的方法来估计单个评级的可靠性在二进制尺度.
- 为了澄清概念关系,并在可重复性和可重复性研究中评估性能.
- 扩展贝叶斯系数的贝叶斯方法,并估计表现尺度可靠性.
主要方法:
- 正常近似,卡帕系数和潜变量方法的比较.
- 对于多重复制的卡帕系数,贝叶斯迪里克莱特多项式方法的扩展.
- 引入贝叶斯方法,从隐性尺度可靠性来估计明显尺度可靠性.
主要成果:
- 正常近似方法的表现不佳.
- 频率主义方法由于奇点问题而表现出不可靠性.
- 隐性变量方法为二进制尺度可靠性提供了一个强大的框架.
结论:
- 对于二进制尺度的可靠性,不建议使用正常近似和频率方法.
- 隐性变量方法提供了一种更可靠的方法来评估二进制结果的可靠性.
- 为心理测量研究中的可靠性估计提供了精细的实际建议.
更多相关视频
09:18Author Spotlight: Assessing the Reliability of Doppler Ultrasound in Measuring Leg Blood Flow
Published on: December 15, 2023
3.4K
09:00Author Spotlight: Validation of SICOLE-R for Assessing Cognitive and Reading Skills in Spanish-Speaking Children and Its Role in Personalized Education
Published on: August 16, 2024
1.2K
相关概念视频
Kendall's Coefficient of Concordance
925
Kendall's Coefficient of Concordance (W), also known as Kendall's W, is a non-parametric statistical measure used to assess the agreement or concordance between multiple raters or judges when they rank a set of items. It is often used when you have ordinal data (ranks) and you want to see if there is consistency or consensus among the raters. It is widely applied in research areas such as psychology, medicine, and social sciences, where multiple judges are asked to rank or rate subjects...
925
Reliability and Validity
13.7K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
13.7K
Kendall's Tau Test
1.1K
Kendall's tau test, also known as the Kendall rank coefficient test, is a nonparametric method for assessing association between two variables. This test is particularly useful for identifying significant correlations when the distributions of the sample and population are unknown. Developed in 1938 by the British statistician Sir Maurice George Kendall, the tau coefficient (denoted as τ) serves as a rank correlation coefficient, with values ranging from -1 to +1.
A τ value of +1 indicates...
A τ value of +1 indicates...
1.1K
Spearman's Rank Correlation Test
1.4K
Spearman's rank correlation test, also known as Spearman's rho, is a nonparametric method for assessing the strength and direction of association between two variables. This test is particularly valuable when the data distribution is unknown or when the assumption of normality does not hold. Named after the English psychologist and statistician Dr. Charles Edward Spearman, it serves as the nonparametric counterpart to Pearson's correlation coefficient.
Spearman's test calculates correlation by...
Spearman's test calculates correlation by...
1.4K
Introduction to Test of Independence
2.9K
In statistics, the term independence means that one can directly obtain the probability of any event involving both variables by multiplying their individual probabilities. Tests of independence are chi-square tests involving the use of a contingency table of observed (data) values.
The test statistic for a test of independence is similar to that of a goodness-of-fit test:
The test statistic for a test of independence is similar to that of a goodness-of-fit test:
2.9K
Ordinal Level of Measurement
31.9K
The way a set of data is measured is called its level of measurement. Correct statistical procedures depend on a researcher being familiar with levels of measurement. For analysis, data are classified into four levels of measurement—nominal, ordinal, interval, and ratio.
Data measured using an ordinal scale are similar to nominal scale data, but there is one major difference. The ordinal scale data can be ordered. An example of ordinal scale data is a list of the top five national parks...
Data measured using an ordinal scale are similar to nominal scale data, but there is one major difference. The ordinal scale data can be ordered. An example of ordinal scale data is a list of the top five national parks...
31.9K
