如果你能复制我:评估个体差异的测量可靠性,跨测量场合和方法的阅读.
Patrick Haller1, Cui Ding1, Maja Stegenwallner-Schütz2,3
1Department of Computational Linguistics, University of Zurich.
Cognitive science
|December 30, 2025
概括
句子处理中的个体差异表明,跨会话和方法的测量不可靠. 在研究心理语言学中的这些个体变异之前,确定测量可靠性至关重要.
科学领域:
- 心理语言学 心理语言学
- 认知科学 认知科学
- 神经科学是一个神经科学.
背景情况:
- 传统的心理语言学理论假设统一的认知机制.
- 最近的研究强调了个人在人类认知和句子处理方面的差异的重要性.
- 可靠性悖论挑战了假设个人层面的影响在会议和方法之间是一致的假设.
研究的目的:
- 评估句子处理中的个体差异的测量可靠性.
- 为了研究效应在多个会话和方法 (眼睛跟踪,自动阅读) 的一致性.
- 检查各种心理语言预测指标的可靠性:单词长度,词汇频率,惊喜,依赖长度和整合负载.
主要方法:
- 收集了一个自然主义的眼动群体,每个参与者有四次会议 (两次眼睛跟踪,两次自动阅读).
- 采用双任务贝叶斯层次模型来分析测量可靠性.
- 对于已确定的心理语言现象,评估了个人层面的影响.
主要成果:
- 对于单词长度效果,跨会话的高可靠性.
- 对词汇频率,依赖距离和集成负载的中等可靠性.
- 对于惊喜来说,可靠性很低.
- 大多数预测器的交叉方法可靠性低至中等,语法集成预测器的可靠性差.
结论:
- 测量可靠性在不同的心理语言预测指标之间有很大的差异.
- 对于更高层次的认知和语法现象,可靠性特别低.
- 确定测量可靠性是关于句子处理中的个人差异的有效推断的关键先决条件.
更多相关视频
06:52Using Cholesky Decomposition to Explore Individual Differences in Longitudinal Relations between Reading Skills
Published on: September 17, 2019
6.7K
06:33Decomposing the Variance in Reading Comprehension to Reveal the Unique and Common Effects of Language and Decoding
Published on: October 11, 2018
7.2K
相关概念视频
Group Design
10.1K
The most basic experimental design involves two groups: the experimental group and the control group. The two groups are designed to be the same except for one difference— experimental manipulation. The experimental group gets the experimental manipulation—that is, the treatment or variable being tested—and the control group does not. Since experimental manipulation is the only difference between the experimental and control groups, we can be sure that any differences between...
10.1K
Reliability and Validity
13.7K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
13.7K
Measures of Intelligence
8.2K
Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
8.2K
Uncertainty in Measurement: Accuracy and Precision
99.2K
Scientists typically make repeated measurements of a quantity to ensure the quality of their findings and to evaluate both the precision and the accuracy of their results. Measurements are said to be precise if they yield very similar results when repeated in the same manner. A measurement is considered accurate if it yields a result that is very close to the true or the accepted value. Precise values agree with each other; accurate values agree with a true value.
99.2K
Random and Systematic Errors
14.3K
Scientists always try their best to record measurements with the utmost accuracy and precision. However, sometimes errors do occur. These errors can be random or systematic. Random errors are observed due to the inconsistency or fluctuation in the measurement process, or variations in the quantity itself that is being measured. Such errors fluctuate from being greater than or less than the true value in repeated measurements. Consider a scientist measuring the length of an earthworm using a...
14.3K
Statistical Analysis: Overview
14.0K
When we take repeated measurements on the same or replicated samples, we will observe inconsistencies in the magnitude. These inconsistencies are called errors. To categorize and characterize these results and their errors, the researcher can use statistical analysis to determine the quality of the measurements and/or suitability of the methods.
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
14.0K
