Related Experiment Video
Updated: Sep 19, 2025

Author Spotlight: Improving Radiation Therapy Access with Radiation Planning Assistant
Published on: October 6, 2023
RadioRAG: Online Retrieval-Augmented Generation for Radiology Question Answering
Soroosh Tayebi Arasteh1,2,3,4, Mahshad Lotfinia1, Keno Bressem5,6
1Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen, Pauwelsstr 30, 52074 Aachen, Germany.
Retrieval-augmented generation (RAG) enhanced large language model (LLM) accuracy in radiology question answering, outperforming human experts for some models. RAG also reduced LLM hallucinations but increased response time.
Area of Science:
- Medical Informatics
- Artificial Intelligence in Radiology
- Natural Language Processing
Background:
- Large Language Models (LLMs) show promise for medical question answering.
- LLMs can generate inaccurate information (hallucinations) without access to real-time data.
- Retrieval-augmented generation (RAG) integrates external knowledge to improve LLM factuality.
Purpose of the Study:
- To assess the diagnostic accuracy of LLMs in radiology question answering.
- To evaluate the impact of Retrieval-Augmented Generation (RAG) on LLM performance.
- To compare LLM performance with and without RAG against human expert accuracy.
Main Methods:
- Developed RadioRAG, a framework for real-time data retrieval from radiologic sources.
- Evaluated multiple LLMs (GPT-3.5-turbo, GPT-4, Mistral, Llama3) with and without RadioRAG.
- Used 80 RSNA Case Collection questions and 24 expert-curated questions for assessment.
Main Results:
- RadioRAG improved accuracy for GPT-3.5-turbo (74% vs 66%) and Mixtral 8x7B (76% vs 65%) on the RSNA-RadioQA dataset.
- LLM accuracy with RAG surpassed human expert performance (63%) for some models.
- RadioRAG reduced hallucinations across all tested LLMs (6%-25% rate) but increased response time fourfold.
Conclusions:
- RadioRAG demonstrates potential to enhance LLM accuracy and factuality in radiology.
- Integrating real-time, domain-specific data via RAG is crucial for reliable AI in medical QA.
- Further research can optimize RAG for faster, more accurate AI-assisted radiology.
More Related Videos
03:14Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
Related Concept Videos
Positron Emission Tomography
One of the main requirements of a PET scan is a positron-emitting radioisotope, which is produced in a cyclotron and then attached to a substance used by the part of the body...
Computed Tomography
The technique was invented in the 1970s and is based on the principle that as X-rays pass through the body, they are absorbed or reflected at different levels. In the technique, a patient lies on a motorized platform while a computerized axial tomography (CAT) scanner rotates...
X-ray Imaging
Radiological Investigation I: X-ray and CT
ER Retrieval Pathway
The ER uses many checkpoints to prevent the entry of incorrectly folded or a resident protein as cargo onto a transport vesicle. These mechanisms...
Radiological Investigation III: Pulmonary Angiogram and PET Scan
Pulmonary Angiogram
A Pulmonary Angiogram is an invasive procedure involving injecting a contrast medium through a catheter threaded into the pulmonary artery or the right side of the heart to visualize the pulmonary vasculature. Computed Tomography (CT) scans have mainly replaced this...