Evaluation Methods for Inference-Time Retrieval-Augmented and Graph Retrieval-Augmented Large Language Models in

Yuhan Zhao1, Yiqun Miao1, Rongrong Guo1

  • 1School of Nursing, Capital Medical University, No. 10 Xitoutiao, Youanmenwai, Fengtai District, Beijing, 100069, China, 86 13910789837.

Summary

Evaluation of retrieval-augmented generation (RAG) and graph-structured RAG (GraphRAG) in healthcare LLMs is inconsistent. Gaps exist in real-world testing, safety, and detailed verification, hindering clinical readiness.

Related Concept Videos