Related Experiment Video
Updated: Jun 17, 2025

Precision Measurements and Parametric Models of Vertebral Endplates
Published on: September 17, 2019
Benchmarking Large Language Models for Cervical Spondylosis
Boyan Zhang1,2, Yueqi Du1,2, Wanru Duan1,2
1Xuanwu Hospital, Capital Medical University, Beijing, China.
Abstract:
Cervical spondylosis is the most common degenerative spinal disorder in modern societies. Patients require a great deal of medical knowledge, and large language models (LLMs) offer patients a novel and convenient tool for accessing medical advice. In this study, we collected the most frequently asked questions by patients with cervical spondylosis in clinical work and internet consultations. The accuracy of the answers provided by LLMs was evaluated and graded by 3 experienced spinal surgeons. Comparative analysis of responses showed that all LLMs could provide satisfactory results, and that among them, GPT-4 had the highest accuracy rate. Variation across each section in all LLMs revealed their ability boundaries and the development direction of artificial intelligence.
More Related Videos
03:14Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
07:44Evaluation of Patients' Posture and Gait Profile After Lumbar Fusion Surgery by Video Rasterstereography and Treadmill Gait Analysis
Published on: March 23, 2019