评估ChatGPT聊天机器人是否符合戒烟公共卫生指南:内容分析
Lorien C Abroms1, Artin Yousefi2, Christina N Wysota3
1Department of Prevention & Community Health, Milken Institute School of Public Health, George Washington University, Washington, DC, United States.
大型语言模型 (LLM) 聊天机器人为戒烟建议提供了不同的可靠性. 虽然一些信息是准确的,但LLM聊天机器人省略了关键细节,有时提供错误信息,特别是关于未经证实的方法.
科学领域:
- 人工智能的人工智能
- 公共卫生 公共卫生
- 数字健康数字健康
背景情况:
- 大型语言模型 (LLM) 聊天机器人为提供戒烟信息提供了潜力.
- 这些人工智能聊天机器人提供的信息的可靠性尚未得到充分证实.
研究的目的:
- 为了评估三个ChatGPT聊天机器人提供的戒烟信息的可靠性:Sarah,BeFreeGPT和BasicGPT.
- 评估是否遵守既定的公共卫生指南和戒烟咨询原则.
主要方法:
- 从谷歌搜索中生成了12个与"如何戒烟"相关的常见查询.
- 分析了聊天机器人的响应,以遵守美国预防服务特别工作组的指导方针和咨询原则.
- 使用两个审稿人的独立编码,与第三个编码器解决的差异.
主要成果:
- 聊天机器人响应的平均依赖率为57.1%的依赖指数项目.
- 与BeFreeGPT (50%) 和BasicGPT (47.8%) 相比,莎拉的坚持率显著更高 (72.2%).
- 错误信息存在于22%的回复中,特别是对于不太依据证据的戒烟方法的查询. 常见的建议包括专业咨询 (80.3%) 和尼古丁替代疗法 (52.7%).
结论:
- 据证据显示,LLM聊天机器人对基于证据的戒烟指南的遵守程度有所变化.
- 人工智能聊天机器人成功提供了一些信息,但经常省略了关键细节,可能提供错误信息.
- 改善LLM聊天机器人指令是必要的,以提高戒烟建议的准确性和完整性.
更多相关视频
09:42Using Continuous Data Tracking Technology to Study Exercise Adherence in Pulmonary Rehabilitation
Published on: November 8, 2013
08:39Generation of Electronic Cigarette Aerosol by a Third-Generation Machine-Vaping Device: Application to Toxicological Studies
Published on: August 25, 2018
相关概念视频
Lifestyle Factors and Health
Benefits of Physical Activity
Physical activity, whether through structured exercise or casual activities like walking, biking, or dancing, is a cornerstone of a...
Chronic Obstructive Pulmonary Disease-V: Management
Smoking Cessation
Drug Dependence
Levels of Health Promotion and Illness Prevention
In primary prevention, actions taken before disease onset prevent the disease from...
Guidelines and Strategies for Safe Computer Charting
Maintain Confidentiality and Security:
Models of Health Promotion and Illness Prevention II
The agent-host-environment model states that disease results...
