Related Experiment Video
Updated: May 16, 2025

Assessing Early Stage Open-Angle Glaucoma in Patients by Isolated-Check Visual Evoked Potential
Published on: May 25, 2020
Chat GPT vs an experienced ophthalmologist: evaluating chatbot writing performance in ophthalmology
Gabriel Katz1,2, Ofira Zloto3,4, Avner Hostovsky1,2
1Faculty of Medical & Health Sciences, Tel Aviv University, Tel Aviv, Israel.
ChatGPT can write scientific ophthalmology introductions comparable to human experts. Experts could not reliably distinguish AI-generated text from human-written introductions, indicating ChatGPT
Area of Science:
- Ophthalmology
- Artificial Intelligence in Medicine
- Scientific Writing
Background:
- The rapid advancement of artificial intelligence (AI) necessitates evaluating its capabilities in specialized scientific fields.
- Assessing AI's role in generating scientific content is crucial for understanding its potential impact on research and publication.
Purpose of the Study:
- To evaluate the proficiency of ChatGPT in drafting scientific introductions for ophthalmology research papers.
- To compare the quality of ChatGPT-generated introductions against those written by experienced ophthalmologists.
Main Methods:
- ChatGPT 4 was prompted to generate introductions for selected ophthalmology papers.
- Ten experienced ophthalmology specialists evaluated introductions written by ChatGPT and human authors, without knowing the source.
- Evaluations focused on metrics including language, data arrangement, factual accuracy, originality, and data currency.
Main Results:
- Ophthalmology specialists correctly identified the source of introductions in only 57.7% of cases.
- No significant differences were found between ChatGPT and human authors across assessed quality metrics.
- Misclassification rates varied by subspecialty, with Oculoplastics having the highest (66.7%) and Retina the lowest (11.1%).
Conclusions:
- ChatGPT demonstrates significant capability in generating scientific paper introductions in ophthalmology.
- AI-generated introductions are statistically indistinguishable from expert-written ones in key quality aspects.
- Further research should explore AI's utility in other manuscript sections and address ethical implications.
More Related Videos
03:14Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
11:12Driving Simulation in the Clinic: Testing Visual Exploratory Behavior in Daily Life Activities in Patients with Visual Field Defects
Published on: September 18, 2012