Evaluating clinical competencies of large language models with a general practice benchmark

Zheqing Li1, Yiying Yang2, Jiping Lang1

  • 1The Sixth Affiliated Hospital of Sun Yat-sen University, Guangzhou, Guangdong, China.

Nature Communications
|April 16, 2026
PubMed
Summary

Current Large Language Models (LLMs) are not ready for autonomous use in general practice. A new benchmark shows LLMs require human oversight for clinical duties.