Related Experiment Video
Updated: May 13, 2025

05:18
Author Spotlight: Improving Radiation Therapy Access with Radiation Planning Assistant
Published on: October 6, 2023
1.2K
Assessing the Quality and Reliability of ChatGPT's Responses to Radiotherapy-Related Patient Queries: Comparative
Ana Grilo1, Catarina Marques2, Maria Corte-Real2
1Research Center for Psychological Science of the Faculty of Psychology, University of Lisbon to CICPSI, Faculdade de Psicologia, Universidade de Lisboa, Av. D. João II, Lote 4.69.01, Parque das Nações, Lisboa, 1990-096, Portugal, 351 964371101.
JMIR Cancer
|April 16, 2025
Summary
GPT-4 provides higher quality radiotherapy information than GPT-3.5, but both AI models present readability challenges for patients seeking cancer treatment details.
Area of Science:
- Oncology
- Artificial Intelligence in Medicine
- Medical Informatics
Background:
- Patients increasingly use online resources for cancer information, but accuracy and readability are often lacking.
- Artificial intelligence (AI) chatbots like ChatGPT offer potential for accessing medical information, including radiotherapy details.
- The quality and reliability of AI-generated radiotherapy information for patients remain unclear, posing risks of misinformation.
Purpose of the Study:
- To evaluate the quality and reliability of ChatGPT's responses to common radiotherapy questions.
- To compare the performance of two ChatGPT versions, GPT-3.5 and GPT-4, in providing radiotherapy information.
Main Methods:
- 40 common radiotherapy questions were posed to GPT-3.5 and GPT-4.
- Radiotherapy experts rated response quality using the General Quality Score (GQS).
- Response consistency (cosine similarity), readability (Flesch scores), and statistical significance were analyzed.
Main Results:
- GPT-4 outperformed GPT-3.5, achieving higher GQS and fewer low-quality ratings.
- Both versions showed high response similarity (median cosine similarity 0.81).
- Readability scores indicated college-level text, challenging for the general public.
Conclusions:
- GPT-4 demonstrates superior capability in answering radiotherapy queries compared to GPT-3.5.
- Both AI models provide information that is difficult for the general public to understand.
- ChatGPT shows promise for patient radiotherapy information but requires strategies to improve readability and mitigate misinformation risks.
Keywords:
ChatGPTOpenAIaccuracyartificial intelligencecancer awarenesschat generative pretrained transformerchatbothealth informationinternet accesslarge language modelnatural language processingpatient informationpatient querypatients with cancerqualityradiotherapyreadability
