Evaluating the reference accuracy of large language models in radiology: a comparative study across subspecialties

Yasin Celal Güneş1, Turay Cesur2, Eren Çamur3

  • 1Kırıkkale Yüksek İhtisas Hospital, Clinic of Radiology, Kırıkkale, Türkiye.

Summary

Claude 3.5 Sonnet excels at generating accurate radiology references, significantly outperforming other large language models like ChatGPT and Google Gemini. This advancement offers dependable citations for radiology research and education, mitigating risks of misinformation.