Related Experiment Video
Updated: Sep 12, 2025

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Constructing a High-Quality English and Korean Medical Corpus for Medical Large Language Model
Chansik Kim1,2, Tong Min Kim2, Youngron Lee2
1Department of Medical Sciences, College of Medicine, The Catholic University of Korea, Seoul, Republic of Korea.
Abstract:
This study presents a high-quality bilingual (English and Korean) medical corpus to support large-scale language models in healthcare. Data from four major hospitals in Korea were converted to Markdown and organized in JSON for efficient use. Evaluation metrics ensured consistency and reliability. The corpus serves as an open, multilingual resource for developing medical LLMs.
More Related Videos
Related Concept Videos
Improving Translational Accuracy
Language Development
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...
Components of Language
One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation
On...

