Related Experiment Video
Updated: Jul 28, 2026

A Knowledge Graph Approach to Elucidate the Role of Organellar Pathways in Disease via Biomedical Reports
Published on: October 13, 2023
Large Language Model-Generated Expansion of the RadLex Ontology: Application to Multinational Datasets of Chest CT
Taehee Lee1, Hyungjin Kim1, Seowoo Lee1
1Department of Radiology, Seoul National University Hospital, Seoul National University College of Medicine, 101 Daehak-ro, Jongno-gu, Seoul, 03080, Korea.
None:
BACKGROUND. RadLex (Radiological Society of North America) is a widely used radiology-specific ontology that standardizes terminology for clinical and research uses. However, the ontology's coverage of clinical radiology reports remains limited due to radiologists' linguistic variation. OBJECTIVE. The purpose of this study was to use a large language model (LLM) to generate an expanded set of lexical variants and synonyms for the RadLex ontology and to evaluate the impact of this expansion on lexical coverage and semantic term recognition using clinical radiology reports. METHODS. This retrospective study used an LLM (Gemini 2.0 Flash Thinking, Google) to generate an expansion (lexical variants [morphologic variants, orthographic variants, and acronyms and abbreviations] and strict semantic synonyms) of the 40,000 RadLex preferred terms, with detailed constraints to ensure semantic alignment. Five datasets of clinical chest CT reports were obtained (two from the study institution in Korea [n = 119,098 and n = 245] and three from public datasets [Spain, n = 5213; Turkey, n = 21,304; United States, n = 19,402]). The same LLM was used to parse the reports into lexicon units (concise text strings representing distinct medical concepts). For each dataset, the lexical coverage rate was automatically computed as a measure of the extent to which the parsed units matched a given expression list. Additionally, 100 randomly selected reports from each dataset were manually reviewed to determine a given expression list's precision, recall, and F1 score (measures of unit-level matching performance when requiring semantic fidelity). Metrics were compared between the existing RadLex-provided expansion and the LLM-generated expansion. RESULTS. The RadLex-provided expansion contained 17,515 terms. The LLM-generated expansion contained 208,465 lexical variants and 69,918 synonyms. For all five datasets, the LLM-generated expansion, compared with the RadLex-provided expansion, had a greater lexical coverage rate (81.9-86.2% vs 67.5-75.4%), greater recall (81.691.4% vs 64.0-80.3%), lower precision (94.8-98.2% vs 100.0% [all datasets]), and greater F1 score (0.91-0.95 vs 0.86-0.91). CONCLUSION. Across multinational datasets of clinical chest CT reports, the LLM-generated term expansion yielded improved lexical coverage and semantic recall, with only small loss of semantic precision, compared with the RadLex-provided expansion. CLINICAL IMPACT. The LLM-based approach provides a practical and scalable solution for expanding radiology ontologies while maintaining semantic alignment; the method can aid real-world natural language processing applications.

