个人差异研究的评分故事回忆:中心细节,外围细节和自动评分
1University of Maryland, College Park, MD, USA. davidmtzphd@gmail.com.
Behavior research methods
|August 7, 2024
概括
使用BERTScore和GPT-4对故事回忆记忆的自动评分对个人差异研究有效. 手动评分可以通过不区分中心和外围细节来简化.
科学领域:
- 认知心理学 认知心理学
- 心理测量 心理测量 心理测量
- 计算语言学 计算语言学
背景情况:
- 故事回忆是一种有价值的情节性记忆范式,但其手动评分的复杂性限制了其在个人差异研究中的使用.
- 差异心理学家经常避免回忆故事,因为得分自由回忆叙事的劳动密集型性质.
研究的目的:
- 确定是否在故事回忆中区分中心和外围细节是个人差异研究所必需的.
- 通过BERTScore和GPT-4等计算方法研究自动化故事回忆得分的可行性.
主要方法:
- 235名参与者完成了故事回忆任务.
- 手动评分的回忆叙事与BERTScore和GPT-4的自动评分进行了比较.
- 分析了评分方法与外部认知因素 (流体智能,结晶智能,工作记忆能力) 之间的相关性.
主要成果:
- 中央和外围细节记忆因素高度相关 (r = .99) 并与外部认知因素类似.
- 来自BERTScore和GPT-4的自动化得分与手动得分有很强的相关性 (r ≥ .97).
- 来自自动化和手动评分方法的因素表现出与外部认知措施相关的类似模式.
结论:
- 故事回忆得分可以通过不考虑细节类型和采用自动得分方法来简化个人差异研究.
- BERTScore 和 GPT-4 显示了自动化故事回忆得分的前景,可能减少手工工作.
- 需要进一步的研究来完善自动评分方法,特别是关于偶尔在计算得分中观察到的瘦身症.
相关概念视频
Group Design
The most basic experimental design involves two groups: the experimental group and the control group. The two groups are designed to be the same except for one difference— experimental manipulation. The experimental group gets the experimental manipulation—that is, the treatment or variable being tested—and the control group does not. Since experimental manipulation is the only difference between the experimental and control groups, we can be sure that any differences between the two are due to...
Review and Preview
In statistics, several tools are used to interpret the data. Measures of central tendency represent the characteristics of the data, such as mean, median, and mode. Additionally, measures of variance like standard deviation and range are used to find the spread of data from the mean. Relative standing measures the distance between data locations. Commonly used measures of relative standings are percentile, z score, and quartiles.
Percentiles are a type of fractile that partition data into...
Percentiles are a type of fractile that partition data into...


