Do ChatGPT and Gemini's Recommendations Align With Established Guidelines for Hand and Upper Extremity Surgery?

Yibin B Zhang1, Fielding S Fischer1, Matthew V Abola2

  • 1Harvard Medical School, Boston, MA, USA.

Hand (New York, N.Y.)
|September 18, 2025
PubMed
Summary

Large language models (LLMs) like ChatGPT and Gemini show moderate clinical accuracy for orthopedic guidelines. Both LLMs have limitations in transparency and consistency, requiring further evaluation before widespread clinical adoption.