A multidimensional benchmarking framework for large language models in oncologic decision making

Mehmet Halici1, Serkan Salturk2, Irem Sayin3

  • 1Department of Radiation Oncology, Basaksehir Cam and Sakura City Hospital, 34480, Istanbul, Turkey. mehmethalici95@gmail.com.

Scientific Reports
|July 21, 2026
PubMed
Summary

Evaluating large language models (LLMs) for oncology clinical decisions requires a multi-dimensional approach. GPT-5 excelled in accuracy and efficiency, demonstrating the value of comprehensive performance scoring for AI tools in cancer care.