第二语言听力测试可靠性的元分析 (1991-2022)
Yuxin Shang1, Vahid Aryadoust1, Zhuohan Hou2
1National Institute of Education, Nanyang Technological University, Singapore 639798, Singapore.
Brain sciences
|August 29, 2024
概括
这一元分析显示,L2听力测试的平均可靠性为0.818,但出版偏差和报告问题可能会影响这些发现. 测试长度和类型显著影响可靠性.
科学领域:
- 获得第二语言的学习.
- 教育测量教育的测量
- 心理测量 心理测量 心理测量
背景情况:
- 评估第二语言 (L2) 听力理解对于语言教学至关重要.
- L2听力测试的可靠性是测量质量的关键指标.
- 以前的研究表明L2听力测试可靠性的变化.
研究的目的:
- 在L2听力测试上进行可靠性概括 (RG) 的元分析.
- 确定影响L2听力评估可靠性的因素.
- 探索报告可靠性估计中的潜在偏见.
主要方法:
- 对92篇发表文章中的122个α系数进行了线性混合效应RG分析.
- 研究基于研究,测试和统计特征的16个变量进行编码.
- 评估了出版偏差和可靠性估计中的异质性.
主要成果:
- 发现L2听力测试的平均可靠性为0.818 (95%CI:0.8030.833).
- 检测到显著的异质性和出版偏差,这表明低可靠性报告不足.
- 项目数量和测试类型 (标准化与非标准化) 是可靠性的重要预测因素.
- 可靠性并没有缓和L2听力得分与其他构造之间的关系.
- 在报告测试结果时观察到可靠性诱导.
结论:
- 虽然L2听力测试显示一般可接受的平均可靠性,但存在显著的异质性和偏差.
- 测试开发人员应考虑项目数和测试设计,以提高可靠性.
- 研究人员和教育工作者需要意识到L2听力测试可靠性的潜在报告偏见和局限性.
相关概念视频
Reliability and Validity
12.7K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
12.7K
Longitudinal Research
11.9K
Sometimes we want to see how people change over time, as in studies of human development and lifespan. When we test the same group of individuals repeatedly over an extended period of time, we are conducting longitudinal research. Longitudinal research is a research design in which data-gathering is administered repeatedly over an extended period of time. For example, we may survey a group of individuals about their dietary habits at age 20, retest them a decade later at age 30, and then again...
11.9K
Measures of Intelligence
7.1K
Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
7.1K


