Multimodal Artificial Intelligence in Tissue Diagnostics: Vision-Language Models and the Future of Computational

Rong Xia1, Brian Isett2, Jie Chen2

  • 1Department of Pathology, New York University, New York, New York.

Summary

Vision-language models (VLMs) integrate visual and text data for computational pathology. This review covers their applications, evaluation, and challenges, highlighting their potential to assist pathologists.

Related Concept Videos