Related Experiment Video
Updated: Jan 24, 2026

Preparing Adherent Cells for X-ray Fluorescence Imaging by Chemical Fixation
Published on: March 12, 2015
FluoroSAM: A Language-promptable Foundation Model for Flexible X-ray Image Segmentation
Benjamin D Killeen1, Liam J Wang1, Blanca Iñígo1
1Johns Hopkins University, Baltimore, MD 21218, USA.
Abstract:
Language promptable X-ray image segmentation would enable greater flexibility for human-in-the-loop workflows in diagnostic and interventional precision medicine. Prior efforts have contributed task-specific models capable of solving problems within a narrow scope, but expanding to broader use requires additional data, annotations, and training time. Recently, language-aligned foundation models (LFMs) - machine learning models trained on large amounts of highly variable image and text data thus enabling broad applicability - have emerged as promising tools for automated image analysis. Existing foundation models for medical image analysis focus on scenarios and modalities where large, richly annotated datasets are available. However, the X-ray imaging modality features highly variable image appearance and applications, from diagnostic chest X-rays to interventional fluoroscopy, with varying availability of data. To pave the way toward an LFM for comprehensive and language-aligned analysis of arbitrary medical X-ray images, we introduce FluoroSAM, a language-promptable variant of the Segment-Anything Model, trained from scratch on 3M synthetic X-ray images from a wide variety of human anatomies, imaging geometries, and viewing angles. These include pseudo-ground truth masks for 128 organ types and 464 tools with associated text descriptions. FluoroSAM is capable of segmenting myriad anatomical structures and tools based on natural language prompts, thanks to the novel incorporation of vector quantization (VQ) of text embeddings in the training process. We demonstrate FluoroSAM's performance quantitatively on real X-ray images and showcase on several applications how FluoroSAM is a key enabler for rich human-machine interaction in the X-ray image acquisition and analysis context. Code is available at https://github.com/arcadelab/fluorosam.
More Related Videos
07:36Eye Tracking During Visually Situated Language Comprehension: Flexibility and Limitations in Uncovering Visual Context Effects
Published on: November 30, 2018
10:39Measurement of X-ray Beam Coherence along Multiple Directions Using 2-D Checkerboard Phase Grating
Published on: October 11, 2016
Related Concept Videos
X-ray Imaging
Language
Corballis and Suddendorf (2007) and Tomasello and Rakoczy (2003) highlight the role of language in...
X-ray Crystallography
Diffraction
Diffraction is the change in the direction of travel experienced by an electromagnetic wave when it encounters a physical barrier whose dimensions are comparable to those of the wavelength of the light. X-rays are electromagnetic radiation with wavelengths about as long as the distance between neighboring...
Physiological Foundation of Stress
Role of the Sympathetic Nervous System
Adrenaline triggers the...
Social Foundations of Self II: The Generalized Other
Theoretical Foundations of Nursing Practice
Theories provide a perspective to assess patients' conditions and organize data and methods. They also assist in analyzing and interpreting information. They represent a...