Real-world performance of open-source large language models in diabetes diagnosis

Shuting Yang1,2,3, Sujie Liu4,5, Yuxi Ma4,5

  • 1National Clinical Research Center for Metabolic Diseases, Key Laboratory of Diabetes Immunology (Central South University), Ministry of Education, and Department of Metabolism and Endocrinology, The Second Xiangya Hospital of Central South University, Changsha, Hunan, China.

Summary

Open-source large language models (LLMs) excel at complex diabetes subtyping but struggle with rule-based diagnoses like diabetic kidney disease. English prompts performed better on Chinese text, suggesting LLMs can aid clinicians but not replace them.