Related Experiment Video
Updated: Sep 12, 2025

A Computer-Based Platform for Aiding Clinicians in Eating Disorder Analysis and Diagnosis
Published on: May 10, 2022
"Can ChatGPT Answer Patient's Questions?": A Preliminary Analysis.
Donghua Tao1, Karl M Kochendorfer2, Tina Griffin1
1Library of the Health Sciences, University of Illinois Chicago, Chicago IL, US.
ChatGPT-4.0 demonstrates strong performance in answering public medical questions, achieving high scores for scientific accuracy and comprehensiveness. However, users should always verify information with healthcare providers for personalized guidance.
Area of Science:
- Medical Informatics
- Artificial Intelligence in Healthcare
- Consumer Health Information
Background:
- Public access to health information via AI tools like ChatGPT is increasing.
- Evaluating the accuracy and reliability of AI-generated medical answers is a significant challenge for non-experts.
- The trustworthiness of ChatGPT's medical advice for the general public remains largely unassessed.
Purpose of the Study:
- To evaluate the scientific accuracy and comprehensiveness of ChatGPT-4.0's responses to medical questions from the public.
- To assess the performance of GPT-4.0 (ChatGPT-4.0) using a dataset of consumer-generated health queries.
- To provide insights into the quality of AI-driven medical information available to non-specialists.
Main Methods:
- Utilized an existing dataset of consumer health questions from the NIH Genetic and Rare Diseases Information Center (GARD).
- Generated 1467 question-answer pairs using the GPT-4-0613 API.
- Randomly selected 100 pairs for evaluation based on Scientific Accuracy and Comprehensiveness using a 0-5 scale.
Main Results:
- ChatGPT-4.0 achieved high scores, with approximately 90% of answers rated 4 or 5 for Scientific Accuracy and 84% for Comprehensiveness.
- Around 7% and 14% of answers received an average score of 3 for Scientific Accuracy and Comprehensiveness, respectively.
- No significant difference in answer quality was observed based on whether the questions followed a specific framework.
Conclusions:
- ChatGPT-4.0 exhibits a high level of accuracy and comprehensiveness in answering public medical queries.
- Despite strong performance, the study underscores the necessity for healthcare consumers to consult healthcare providers for verification and personalized advice.
- Further research is recommended to explore question phrasing impact and comparative evaluations by consumers versus professionals.
More Related Videos
Related Concept Videos
Techniques of Therapeutic Communication II: Focusing, Paraphrasing, and Summarizing
This therapeutic technique can also be used when a patient brings up pertinent information during a health-related conversation. The...
Assessment of the Gastrointestinal System I: Subjective Data
Health History
The initial step in assessing the GI system is obtaining a comprehensive health history. This includes inquiring about the patient's history or presence of problems...
Patient-centered Care
Assessment of the Gastrointestinal System II: Health Perception Pattern
Health Perception Patterns
Health perception patterns offer valuable insights into a patient's lifestyle habits and how they may impact their GI health. These patterns include:
SBAR II: Application of SBAR
SBAR Report from a Nurse to a Health Care Provider
S: "Hello, Dr. Smith. This is Jane, RN, from the Med Surg unit. I am calling to tell you about Ms. White in Room 210, who is experiencing increased pain and redness at her incision site. Her recent...
Barriers to Effective Communication II
Cultural barriers:
Differences in values, beliefs, religion, knowledge, and tradition can significantly impact communication. Awareness of nonverbal cues is critical, especially when conversing with a patient from a different culture. What appears appropriate in one culture may be inappropriate in another.
Semantic barriers:
As a result of their tendency to use...

