Adversarial prompt and fine-tuning attacks threaten medical large language models

Yifan Yang1,2, Qiao Jin1, Furong Huang2

  • 1National Library of Medicine (NLM), National Institutes of Health (NIH), Bethesda, MD, USA.

Nature Communications
|October 9, 2025
PubMed
Abstract