Related Experiment Video
Updated: May 10, 2025

Eye-Tracking Control to Assess Cognitive Functions in Patients with Amyotrophic Lateral Sclerosis
Published on: October 13, 2016
Can OpenAI's New o1 Model Outperform Its Predecessors in Common Eye Care Queries?
Krithi Pushpanathan1,2, Minjie Zou1,2, Sahana Srinivasan1,2
1Department of Ophthalmology, Yong Loo Lin School of Medicine, National University of Singapore, Singapore.
OpenAI's o1 model shows improved correctness and readability in eye care queries compared to ChatGPT-4, but performs similarly to ChatGPT-4o. While o1 excels in some areas, its completeness for ocular symptoms needs improvement.
Area of Science:
- Ophthalmology
- Artificial Intelligence in Healthcare
- Medical Informatics
Background:
- The rapid advancement of AI models like OpenAI's o1 promises enhanced reasoning for specialized queries.
- Assessing the performance of new AI models in critical fields such as eye care is essential.
- Previous AI models have shown variable performance in responding to ophthalmic questions.
Purpose of the Study:
- To evaluate and compare the performance of OpenAI's o1, ChatGPT-4o, and ChatGPT-4 in answering ophthalmic-related queries.
- To assess AI models based on correctness, completeness, and readability in the context of eye care.
- To identify strengths and weaknesses of o1 in addressing specific ophthalmic subtopics.
Main Methods:
- A cross-sectional study design was employed.
- Sixteen challenging ophthalmic queries (myopia, ocular symptoms, retinal conditions) were used.
- Responses were evaluated by three masked ophthalmologists on correctness, completeness, and readability (5-point scale).
Main Results:
- OpenAI's o1 achieved the highest scores for correctness (12.6) and readability (14.2), outperforming ChatGPT-4.
- No significant performance difference was observed between o1 and ChatGPT-4o.
- ChatGPT-4o led in completeness (12.4), with o1 scoring 10.8; o1 showed limitations in completeness for ocular symptom queries (5.5/15).
Conclusions:
- OpenAI's o1 demonstrates superior correctness and readability over ChatGPT-4 for ophthalmic queries, aligning with ChatGPT-4o's performance.
- Despite marketing claims, o1's overall performance in eye care queries is comparable to ChatGPT-4o.
- Further improvements in completeness, particularly for specific conditions like ocular symptoms, are needed for o1.
More Related Videos
05:49Author Spotlight: Deciphering Electrical Networks Behind Complex Brain Activities and Disorders
Published on: November 1, 2024
07:12Development of a Gaze-Contingent Display Framework Designed for Perceptual and Oculomotor Research with Simulated Central Vision Loss
Published on: April 11, 2025
Related Concept Videos
Accessory Structures of the Eye
Open Angle Glaucoma: Treatment
Drugs such as carbonic anhydrase inhibitors, α2- and...
Glaucoma: Overview
Angle Closure Glaucoma: Treatment
Muscles of the Eye
Extraocular Muscles
The six extraocular muscles surround the eyeball and control its movements. They are responsible for a wide range of eye motions, including looking up, down, left, right, and...