Related Experiment Video
Updated: Aug 20, 2025

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Keyword-augmented and semi-automatic generation of FESS reports: a proof-of-concept study
V Kunz1, V Wildfeuer2, R Bieck3
1Department of Otolaryngology, Head and Neck Surgery, University Hospital Leipzig, Liebigstraße 10-14, 04103, Leipzig, Germany. Viktor.Kunz@medizin.uni-leipzig.de.
Introduction:
Surgical reports are usually written after a procedure and must often be reproduced from memory. Thus, this is an error-prone, and time-consuming task which increases the workload of physicians. In this proof-of-concept study, we developed and evaluated a software tool using Artificial Intelligence (AI) for semi-automatic intraoperative generation of surgical reports for functional endoscopic sinus surgery (FESS).
Materials And Methods:
A vocabulary of keywords for developing a neural language model was created. With an encoder-decoder-architecture, artificially coherent sentence structures, as they would be expected in general operation reports, were generated. A first set of 48 conventional operation reports were used for model training. After training, the reports were generated again and compared to those before training. Established metrics were used to measure optimization of the model objectively. A cohort of 16 physicians corrected and evaluated three randomly selected, generated reports in four categories: "quality of the generated operation reports," "time-saving," "clinical benefits" and "comparison with the conventional reports." The corrections of the generated reports were counted and categorized.
Results:
Objective parameters showed improvement in performance after training the language model (p < 0.001). 27.78% estimated a timesaving of 1-15 and 61.11% of 16-30 min per day. 66.66% claimed to see a clinical benefit and 61.11% a relevant workload reduction. Similarity in content between generated and conventional reports was seen by 33.33%, similarity in form by 27.78%. 66.67% would use this tool in the future. An average of 23.25 ± 12.5 corrections was needed for a subjectively appropriate surgery report.
Conclusion:
The results indicate existing limitations of applying deep learning to text generation of operation reports and show a high acceptance by the physicians. By taking over this time-consuming task, the tool could reduce workload, optimize clinical workflows and improve the quality of patient care. Further training of the language model is needed.
Related Concept Videos
SBAR II: Application of SBAR
SBAR Report from a Nurse to a Health Care Provider
S: "Hello, Dr. Smith. This is Jane, RN, from the Med Surg unit. I am calling to tell you about Ms. White in Room 210, who is experiencing increased pain and redness at her incision site. Her recent...
Reporter Genes
Non-equilibrium in the Cell
Types of Reports II: Incident or Occurrence Report
Purposes:
In the healthcare industry, reports play a crucial role in documenting incidents within an agency. The primary objective of these reports is to ensure patient safety, uphold the...
Generation of Straight or Branched Actin Filaments
Arp2/3 Complex
Arp2/3 complex is a seven-subunit complex consisting of two proteins similar to actin- Arp2 and Arp3, and five other subunits that help keep Arp2 and Arp3 inactive. When required, the complex is...

![An Automated Radiosynthesis of [68Ga]Ga-FAPI-46 for Routine Clinical Use](/_next/image?url=https%3A%2F%2Fcloudfront.jove.com%2FCDNSource%2Fteasers%2F66708.jpg&w=3840&q=50)