Related Experiment Video
Updated: Jun 5, 2025

10:42
A Postoperative Evaluation Guideline for Computer-Assisted Reconstruction of the Mandible
Published on: January 28, 2020
6.5K
Chatbots in Limb Lengthening and Reconstruction Surgery: How Accurate Are the Responses?
Anirejuoritse Bafor1, Daryn Strub1, Søren Kold2
1Department of Orthopedic Surgery Nationwide Children's Hospital, Columbus, OH.
Journal of Pediatric Orthopedics
|December 11, 2024
Summary
ChatGPT demonstrated superior accuracy in answering limb reconstruction surgery questions compared to Google Bard and Microsoft Copilot. While ChatGPT
Area of Science:
- Orthopaedic Surgery
- Artificial Intelligence in Medicine
Background:
- Healthcare information seeking is increasingly reliant on AI chatbots.
- Parents frequently use online resources for pediatric health concerns.
- AI chatbot accuracy in limb reconstruction surgery remains unevaluated.
Purpose of the Study:
- To evaluate the accuracy of AI chatbots for limb reconstruction surgery information.
- Compare ChatGPT, Google Bard, and Microsoft Copilot response accuracy.
Main Methods:
- Generated 23 common limb reconstruction surgery questions.
- Posed questions to 3 AI chatbots on 3 separate occasions.
- Orthopaedic surgeons rated randomized, blinded responses using a 4-point scale.
Main Results:
- ChatGPT achieved the highest response accuracy score.
- Microsoft Copilot demonstrated the lowest accuracy.
- Findings were consistent across all three orthopaedic surgeon raters.
Conclusions:
- ChatGPT responses were deemed satisfactory with minimal clarification needed.
- Microsoft Copilot responses required moderate clarification.
- AI chatbot accuracy varies significantly in specialized medical fields.

