Automating methodological quality assessment in orthopedic systematic reviews using large language models.

Yu-Jui Huang1, Kai-Cheng Chang2, Ying-Chen Kuo3

  • 1Department of Orthopedic Surgery, Linkou Chang Gung Memorial Hospital, Taoyuan City, Taiwan.

Summary

This study evaluated whether artificial intelligence tools could accurately assess the quality of orthopedic research papers. Researchers compared automated ratings from three different language models against human expert evaluations. While the technology showed high agreement with human reviewers, it struggled with complex tasks requiring subjective judgment. The authors suggest that these tools are best used as assistants to improve efficiency rather than as replacements for human experts.

Frequently Asked Questions