Related Experiment Video
Updated: Sep 12, 2025

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Towards Sustainable Inference of LLMs for Medical Education Through Token Count Minimization
Hyunggu Jung1,2, Jiyoo Min3, Yunseo Moon3
1College of Nursing, Seoul National University, Seoul, Republic of Korea.
Abstract:
The carbon emissions generated during the inference phase of large language models (LLMs) are a growing concern. We propose a method to minimize the number of tokens of input prompts for medical students interacting with LLMs. Using English physical examination course materials, we applied translation and paraphrasing to recommend the most token-efficient prompt. We found that English baseline prompts had fewer tokens than Korean ones, while the paraphrased forms we proposed significantly reduced token counts compared to the baseline prompts.
More Related Videos
Related Concept Videos
Improving Translational Accuracy
Language and Cognition
Language Development
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...

