Related Experiment Video
Updated: Jun 18, 2026

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Routing distilled knowledge via mixture of LoRA experts for large language model based bundle generation
Kaidong Feng1, Zhu Sun2, Hui Fang3
1School of Computer Science and Engineering, Yanshan University, Qinhuangdao, 066004, Hebei, China.
Abstract:
Large Language Models (LLMs) have shown potential in automatic bundle generation but suffer from prohibitive computational costs. Although knowledge distillation offers a pathway to more efficient student models, our preliminary study reveals that naively integrating diverse types of distilled knowledge from teacher LLMs into student LLMs leads to knowledge conflict, negatively impacting the performance of bundle generation. To address this, we propose RouteDK, a framework for routing distilled knowledge through a mixture of Low-Rank Adaptation (LoRA) expert architecture. Specifically, we first distill two complementary types of knowledge from the teacher LLM: high-level knowledge (generalizable rules) and fine-grained knowledge (session-specific reasoning). We then train knowledge-specific LoRA experts for each type of knowledge together with a base LoRA expert. For effective integration, we propose a dynamic fusion module, featuring an input-aware router, where the router balances expert contributions by dynamically determining optimal weights based on input, thereby effectively mitigating knowledge conflicts. To further improve inference reliability, we design an inference-time enhancement module to reduce variance and mitigate suboptimal reasoning. Experiments on three public datasets show that our RouteDK achieves accuracy comparable to or even better than the teacher LLM, while maintaining strong computational efficiency. In addition, it outperforms state-of-the-art approaches for bundle generation.
Related Concept Videos
Language Development
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...
Language
Corballis and Suddendorf (2007) and Tomasello and Rakoczy (2003) highlight the role of language in...
Distribution Reliability and Automation
Components of Language
Theories of Dissolution: Diffusion Layer Model
This process starts with a thin layer, saturated with the drug, forming at the interface between the solid and liquid. The solute then diffuses from this layer into the main solution. The Noyes-Whitney equation suggests that the rate of dissolution relies on the diffusion...
Improving Translational Accuracy