Language-Guided Segmentation of Medical Images: A Review of Foundation Models

Saqib Qamar1,2

  • 1Department of Intelligent Systems, KTH Royal Institute of Technology, 10044 Stockholm, Sweden.

Summary

Vision-language foundation models revolutionize medical image segmentation using text prompts for diverse tasks. This survey explores their technical background, taxonomy, adaptation, and clinical applications, highlighting challenges and future directions.

Related Concept Videos