The In-depth Comparative Analysis of Four Large Language AI Models for Risk Assessment and Information Retrieval from

Lun-Hsiang Yuan1,2, Shi-Wei Huang2,3, Dean Chou1,4,5,6

  • 1Department of Biomedical Engineering, National Cheng-Kung University, Tainan, Taiwan.

PubMed
Summary

Four large language models (LLMs) were evaluated for information retrieval and risk assessment in prostate cancer (PC) reports. ChatGPT-4-turbo showed the highest accuracy in risk assessment, indicating potential for clinical decision support.