生成型人工智能的作用在评估遵守负责任的新闻媒体关于自杀的报道中的作用:一个多站点,三种语言的研究研究
Zohar Elyospeh1, Bénédicte Nobile2,3, Inbar Levkovich4
1https://ror.org/02f009v59University of Haifa, Mount Carmel, Haifa, Israel.
概括
人工智能 (AI) 模型在评估媒体遵守世界卫生组织 (WHO) 自杀报告指南时,与人类判断有很强的一致性. 这些人工智能工具可以帮助记者促进负责任的报道,并支持全球预防自杀的努力.
科学领域:
- 公共卫生 公共卫生
- 人工智能的人工智能
- 媒体研究 媒体研究
背景情况:
- 媒体遵守世界卫生组织 (WHO) 的指导方针对于预防自杀至关重要.
- 需要有效,快速的方法来评估这种坚持.
研究的目的:
- 评估人工智能 (AI) 模型评估媒体遵守世卫组织自杀报告指南的能力.
- 为了比较Claude Opus 3和GPT-4O与人类评分器的性能.
主要方法:
- 一项比较有效性的研究使用了120篇与自杀相关的英语,希伯来语和法语文章.
- 根据世卫组织的指导方针,六名人类评价者评估了文章.
- 两个人工智能模型 (克劳德·奥普斯3,GPT-4O) 评估了文章的坚持性.
- 使用类内相关系数 (ICC) 和斯皮尔曼相关性来衡量协议.
主要成果:
- 在所有语言中,对世卫组织指南的总体遵守率约为50%.
- 两种人工智能模型都显示出与人类评级者有很强的一致性.
- GPT-4O显示了最高的个人协议 (ICC = 0.793),并结合AI评估显示了最高的可靠性 (ICC = 0.812).
结论:
- 人工智能模型可以有效地复制人类判断,评估媒体是否遵守世卫组织自杀报告指南.
- 虽然人工智能工具有望提高负责任的报告和预防自杀,但人类监督仍然至关重要.
- 人工智能有可能通过促进更好的新闻实践来支持全球预防自杀的努力.
相关概念视频
Surveys
15.4K
Often, psychologists develop surveys as a means of gathering data. Surveys are lists of questions to be answered by research participants, and can be delivered as paper-and-pencil questionnaires, administered electronically, or conducted verbally. Generally, the survey itself can be completed in a short time, and the ease of administering a survey makes it easy to collect data from a large number of people.
15.4K
Stereotype Content Model
14.9K
The Stereotype Content Model (SCM) was first proposed by Susan Fiske and her colleagues (Fiske, Cuddy, Glick & Xu, 2002; see also Fiske, 2012 and Fiske, 2017). The SCM specifies that when someone encounters a new group, they will stereotype them based on two metrics: warmth—or that group’s perceived intent, and how likely they are to provide help or inflict harm—and competence—or their ability to carry out that objective. Depending on the warmth-competence...
14.9K


