Evaluation of large language models in female malignancy Q&A

Kai Xin1, Sen Hong2, Xiahui Wu3

  • 1Department of Oncology, Nanjing Drum Tower Hospital, Affiliated Hospital of Medical School, Nanjing University, Nanjing, China.

Medicine
|September 17, 2025
PubMed
Summary

Large language models (LLMs) show potential in answering questions about female malignancy, with Llama-3.1-405B, OpenAI o1, and DeepSeek-R1 performing best. Challenges remain with complex medical terminology and flexible scenarios.