对于密集的纵向多级数据的非图形测量器间可靠性测量.
Tobias Koch1, Miriam F Jaehne1, Michaela Riediger1
1Friedrich-Schiller-Universität Jena, Jena, Germany.
The British journal of mathematical and statistical psychology
|December 20, 2025
概括
这项研究引入了一种新的模型,用于评估心理研究中不同类型的评级者之间的协议. 关系的持续时间和合作伙伴的认知资源提高了对创新的评级者一致性.
科学领域:
- 心理学 心理学 心理学
- 量化心理学 量化心理学
- 心理测量 心理测量 心理测量
背景情况:
- 在心理学研究中,评分器间的可靠性是必不可少的.
- 现有的模型经常与结构上不同的评级类型 (例如,自我与合作伙伴报告) 斗争.
- 需要方法来评估个人特定的评级者一致性及其预测因素.
研究的目的:
- 提出一种新的多层次潜伏时间序列模型 (MR-MLTS),用于具有多个结构不同评分器的强度纵向数据.
- 能够估计个人特异性 (个人特异性) 评级者一致性系数.
- 促进将评级者的一致性与外部变量联系起来.
主要方法:
- 开发了一种多层潜伏时间序列模型 (MR-MLTS).
- 在Mplus和一个新的R包 (mlts) 中实现了模型.
- 将模型应用于100对异性恋夫妇 (86个时间点) 的密集纵向数据.
主要成果:
- 关系的持续时间和合作伙伴的认知资源积极预测评级者对创新的一致性.
- 模拟结果突出了时间点数在估计特征系数时的重要性.
- 参与者数量对于恢复随机效应差异至关重要.
结论:
- 该MR-MLTS模型提供了一种灵活的方法来量化复杂的二度和纵向研究中的评级者协议.
- 调查结果强调了关系因素在塑造评级者一致性的作用.
- 该模型提供了有关测量精度和不同协议的个体差异的宝贵见解.
相关概念视频
Reliability and Validity
13.7K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
13.7K
Statistical Analysis: Overview
14.0K
When we take repeated measurements on the same or replicated samples, we will observe inconsistencies in the magnitude. These inconsistencies are called errors. To categorize and characterize these results and their errors, the researcher can use statistical analysis to determine the quality of the measurements and/or suitability of the methods.
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
14.0K
Kendall's Coefficient of Concordance
915
Kendall's Coefficient of Concordance (W), also known as Kendall's W, is a non-parametric statistical measure used to assess the agreement or concordance between multiple raters or judges when they rank a set of items. It is often used when you have ordinal data (ranks) and you want to see if there is consistency or consensus among the raters. It is widely applied in research areas such as psychology, medicine, and social sciences, where multiple judges are asked to rank or rate subjects...
915
Group Design
10.1K
The most basic experimental design involves two groups: the experimental group and the control group. The two groups are designed to be the same except for one difference— experimental manipulation. The experimental group gets the experimental manipulation—that is, the treatment or variable being tested—and the control group does not. Since experimental manipulation is the only difference between the experimental and control groups, we can be sure that any differences between...
10.1K
Bioequivalence Experimental Study Designs: Repeated Measures, Cross-Over, Carry-Over, and Latin Square Designs
162
Body:Bioequivalence experimental study designs play a pivotal role in testing the effectiveness of various treatments. Key among these are the repeated measures, cross-over, carry-over, and Latin square designs. In the repeated measures design, each subject receives all treatments, allowing for temporal comparisons. This type of design is useful in reducing variability but requires careful planning to avoid bias.The cross-over design, an economical method, involves sequential administration of...
162
Longitudinal Research
13.0K
Sometimes we want to see how people change over time, as in studies of human development and lifespan. When we test the same group of individuals repeatedly over an extended period of time, we are conducting longitudinal research. Longitudinal research is a research design in which data-gathering is administered repeatedly over an extended period of time. For example, we may survey a group of individuals about their dietary habits at age 20, retest them a decade later at age 30, and then again...
13.0K


