Related Experiment Video
Updated: Apr 30, 2026

Assessment and Communication for People with Disorders of Consciousness
Published on: August 1, 2017
Comparative evaluation of ChatGPT and Gemini in brain-computer interfaces patient education: A multi-dimensional
Shichao Liu1, Lingning Su1, Qiuyu He1
1Department of Neurosurgery, Fujian Medical University Union Hospital, Fuzhou, Fujian 350001, China.
Background:
Brain-Computer Interfaces (BCI) are a type of life-altering neurotechnology, but their inherent complexity poses significant challenges to patient education. Large Language Models (LLMs), such as ChatGPT and Gemini, offer new possibilities to address this challenge. This study aims to conduct a multi-dimensional, rigorous comparative analysis of the performance of these two mainstream AI models in responding to common patient questions related to BCI.
Methods:
Through a structured process combining clinical expert consensus, literature review, and online patient community analysis, we identified 13 key patient questions covering the entire BCI treatment cycle. We then obtained responses to these questions from ChatGPT and Gemini on September 1, 2025. An evaluation panel, composed of clinical experts and non-medical professionals, conducted a blinded assessment of the response quality using standardized Likert scales across three dimensions: reliability, accuracy, and comprehensibility. Concurrently, we performed an objective, quantitative analysis of the response texts using the Flesch-Kincaid readability tests.
Results:
On core quality metrics such as reliability, accuracy, and comprehensibility, the performance of the two models was generally comparable, both demonstrating a high level of proficiency with only sporadic statistical differences on a few technical questions. However, a clear significant disparity emerged in the dimension of readability: for 12 of the 13 questions, the text generated by Gemini required a significantly lower reading grade level than that of ChatGPT (p < 0.05) and had significantly higher reading ease scores. This difference stemmed from Gemini's tendency to use shorter sentences and simpler vocabulary.
Conclusion:
AI chatbots possess immense potential in the field of BCI patient education. Although both ChatGPT and Gemini can provide high-quality information, Gemini demonstrates a clear advantage in the accessibility and approachability of information, making it a potentially more suitable tool for initial application across diverse patient populations. Nevertheless, the limitations of AI in handling highly specialized and dynamically changing knowledge underscore the indispensable role of human expert supervision and validation in any clinical application.

