Benchmarking the diagnostic performance of open source LLMs in 1933 Eurorad case reports

Su Hwan Kim1, Severin Schramm2, Lisa C Adams3

  • 1Department of Diagnostic and Interventional Neuroradiology, Klinikum rechts der Isar, School of Medicine and Health, Technical University of Munich, Munich, Germany. suhwan.kim@tum.de.

NPJ Digital Medicine
|February 11, 2025
PubMed
Summary

Open-source large language models (LLMs) show promise for supporting radiological diagnostics, with Llama-3-70B closely matching proprietary models like GPT-4o in performance. These AI tools can aid in differential diagnosis for complex cases.