赋予患者权力:在癌教育中,大语言模型的准确性和可读性如何?
Abdulghafour Halawani1, Sultan G Almehmadi1, Bandar A Alhubaishy1
1Department of Urology, King Abdulaziz University, Jeddah, Saudi Arabia.
Frontiers in oncology
|October 11, 2024
概括
人工智能生成的癌信息显示高准确性,但可以省略细节. 虽然聊天机器人可以简化内容,但人工智能和专家材料都可能超过一般可读性水平,需要医生仔细指导.
科学领域:
- 医疗信息学 医疗信息学
- 医疗保健中的人工智能
- 患者教育 患者教育
背景情况:
- 人工智能 (AI) 正在改变医疗保健,特别是在创建个性化的患者教育材料 (PEM) 方面.
- 评估人工智能生成的癌信息对于指导医生和患者至关重要.
- 这项研究将AI工具 (ChatGPT 4.0,Gemini,Perplexity) 与已建立的泌尿病学协会PEM进行比较.
研究的目的:
- 评估人工智能生成的癌信息的准确性和可读性.
- 将AI输出与美国泌尿器官学会 (AUA) 和欧洲泌尿器官学会 (EAU) 的PEM进行比较.
- 为医生提供有关利用人工智能资源用于患者教育的指导.
主要方法:
- 从AUA和EAU收集和分类PEM.
- 使用谷歌趋势来识别癌查询用于AI输入.
- 评估AI输出准确度,使用4名审稿人的5分利克特级别.
- 通过枪雾指数,SMOG和Flesch-Kincaid等级公式评估可读性.
- 指示AI聊天机器人将内容简化到六年级的阅读水平.
主要成果:
- AUA PEM是最易读的 (9.84 ± 1.2),其次是双子AI (10.83 ± 2.31),聊天GPT-4.0 (11.03 ± 1.76),EAU (11.88 ± 1.11),以及困惑AI (12.66 ± 1.83).
- 根据要求,AI聊天机器人成功地将文本简化为较低的年级.
- 人工智能生成的内容显示了高的整体准确性,其中有轻微的遗漏和一些不准确性,特别是在治疗信息中.
- 权威的PEM和AI输出都经常超过了推的可读性水平.
结论:
- 虽然AUA PEM是最易读的,但专家和AI材料可能不符合一般人群的可读性标准.
- 人工智能聊天机器人可以在提示时调整内容的复杂性,但准确性可能会受到细节遗漏和不准确的影响.
- 人工智能工具在患者教育中显示出作为辅助资源的潜力,但由于性能变化,需要谨慎应用.
更多相关视频
相关概念视频
Cancer Survival Analysis
328
Cancer survival analysis focuses on quantifying and interpreting the time from a key starting point, such as diagnosis or the initiation of treatment, to a specific endpoint, such as remission or death. This analysis provides critical insights into treatment effectiveness and factors that influence patient outcomes, helping to shape clinical decisions and guide prognostic evaluations. A cornerstone of oncology research, survival analysis tackles the challenges of skewed, non-normally...
328
Mouse Models of Cancer Study
5.5K
Mice have long served as models for studying human biology and pathology because of their phylogenetic and physiological similarity with humans. They are also easy to maintain and breed in the laboratory, and hence, many inbred strains are now available for research. Studies on mice have contributed immeasurably to our understanding of cancer biology.
The development of transgenic, knockout, and knock-in mice has led to an exponential increase in their use as model organisms in research,...
The development of transgenic, knockout, and knock-in mice has led to an exponential increase in their use as model organisms in research,...
5.5K
Improving Translational Accuracy
9.3K
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
9.3K


