测量可变测试长度的可靠性理论,用ERN和Pe在侧面任务中收集的图示来说明
Jules L Ellis1, Klaas Sijtsma2, Kristel de Groot3
1Open University of the Netherlands.
Psychometrika
|February 25, 2026
概括
估计与事件相关的潜在可靠性是具有不平等数据的挑战. 与一般化理论相比,一种新的经典测试理论方法提供了一种更简单,更有洞察力的方法.
科学领域:
- 心理生理学 心理生理学
- 神经科学是一个神经科学.
- 心理测量 心理测量 心理测量
背景情况:
- 在心理生理学中,估计事件相关潜能 (ERP) 的可靠性至关重要.
- 埃里克森侧面任务往往产生不平等数量的观测每人由于响应依赖的神经反应.
- 像概括性理论和贝叶斯估计这样的现有方法存在复杂性.
研究的目的:
- 引入和评估一种使用经典测试理论和频率估计来估计ERP可靠性的新方法.
- 为了应对可靠性分析中每个受试者的不平等观察的挑战.
- 为现有方法提供一种更简单,更有洞察力的替代方案.
主要方法:
- 应用经典测试理论原则到ERP数据.
- 使用频率估计技术进行可靠性评估.
- 开发一个新的整体可靠性系数,用于观察不平等的群体.
主要成果:
- 提出的经典测试理论方法有效地估计了ERP的可靠性.
- 这种方法提供了一个比一般化理论更简单的框架.
- 该方法为以前未解决的可靠性方面提供了额外的见解.
- 对不平等观察的受试者定义了一个新的可靠性系数.
结论:
- 经典测试理论为估计ERP可靠性提供了一种可行且更简单的替代方案,特别是在不平等的观察的情况下.
- 新的频率主义方法提供了有价值的见解和统一的可靠性系数.
- 虽然概括性理论有其优点,但经典方法为特定的心理生理学研究问题提供了引人注目的替代方案.
相关概念视频
Reliability and Validity
14.2K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
14.2K
Uncertainty in Measurement: Accuracy and Precision
111.6K
Scientists typically make repeated measurements of a quantity to ensure the quality of their findings and to evaluate both the precision and the accuracy of their results. Measurements are said to be precise if they yield very similar results when repeated in the same manner. A measurement is considered accurate if it yields a result that is very close to the true or the accepted value. Precise values agree with each other; accurate values agree with a true value.
111.6K
Random and Systematic Errors
15.5K
Scientists always try their best to record measurements with the utmost accuracy and precision. However, sometimes errors do occur. These errors can be random or systematic. Random errors are observed due to the inconsistency or fluctuation in the measurement process, or variations in the quantity itself that is being measured. Such errors fluctuate from being greater than or less than the true value in repeated measurements. Consider a scientist measuring the length of an earthworm using a...
15.5K
Longitudinal Research
13.5K
Sometimes we want to see how people change over time, as in studies of human development and lifespan. When we test the same group of individuals repeatedly over an extended period of time, we are conducting longitudinal research. Longitudinal research is a research design in which data-gathering is administered repeatedly over an extended period of time. For example, we may survey a group of individuals about their dietary habits at age 20, retest them a decade later at age 30, and then again...
13.5K
McNemar's Test
921
McNemar's Test is a nonparametric statistical test used to determine if there is a significant difference in proportions between two related groups when the outcome is binary (e.g., yes/no, success/failure). It is beneficial when we have paired data, such as pre-test/post-test designs, where the same subjects are measured under two different conditions. The test is named after the statistician Quinn McNemar, who introduced it in 1947. It is commonly used in situations where subjects are...
921
Behrens–Fisher Test
293
The Behrens-Fisher test is a statistical method designed to address the Behrens-Fisher problem, which arises when comparing the means of two normally distributed populations with unequal variances. Unlike the Student's t-test, which assumes equal variances, the Behrens-Fisher test allows for mean comparison without this restrictive assumption. This flexibility makes it particularly valuable in scenarios where two independent samples exhibit normality but lack variance homogeneity.
This test...
This test...
293


