Yifan Yang1,2, Qiao Jin1, Furong Huang2

  • 1National Library of Medicine (NLM), National Institutes of Health (NIH), Bethesda, MD, USA.

Nature communications
|October 9, 2025
PubMed
概括

医疗保健中的大型语言模型 (LLM) 容易受到诸如即时注射和中毒数据之类的对抗性攻击. 在攻击后检测模型重量的变化是开发安全医疗AI防御的关键.