Related Experiment Video
Updated: May 8, 2026

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
A methodology for extending domain coverage in SemRep
Graciela Rosemblat1, Dongwook Shin, Halil Kilicoglu
1National Library of Medicine, National Institutes of Health, Lister Hill Center, Cognitive Science Branch, 8600 Rockville Pike, Bethesda, MD 20894, USA.
Abstract:
We describe a domain-independent methodology to extend SemRep coverage beyond the biomedical domain. SemRep, a natural language processing application originally designed for biomedical texts, uses the knowledge sources provided by the Unified Medical Language System (UMLS©). Ontological and terminological extensions to the system are needed in order to support other areas of knowledge. We extended SemRep's application by developing a semantic representation of a previously unsupported domain. This was achieved by adapting well-known ontology engineering phases and integrating them with the UMLS knowledge sources on which SemRep crucially depends. While the process to extend SemRep coverage has been successfully applied in earlier projects, this paper presents in detail the step-wise approach we followed and the mechanisms implemented. A case study in the field of medical informatics illustrates how the ontology engineering phases have been adapted for optimal integration with the UMLS. We provide qualitative and quantitative results, which indicate the validity and usefulness of our methodology.
Related Concept Videos
Methods of Medium Optimization
Method of Sections: Problem Solving II
Method of Sections: Problem Solving I
Extraction: Advanced Methods
Conservation of Protein Domains
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to form...
Field Procedure for Staking Out Curves