Related Experiment Video
Updated: May 2, 2026

Mining Spatial Transcriptomics Datasets using DeepSpaceDB
Published on: September 5, 2025
SpatialFinder: a human-in-the-loop vision-language framework for prioritizing high-value regions in spatial
Jonathan Xu1, Michelle Jiang2, Shunsuke Koga3,4
1The Wharton School, University of Pennsylvania, Philadelphia, PA, United States.
None:
Sequencing an entire spatial transcriptomics slide can cost thousands of dollars per assay, making routine use impractical. Focusing on smaller regions of interest (ROIs) based on adjacent H&E slides offers a practical alternative, but there is (i) no reliable way to identify the most informative areas from standard H&E images alone; and (ii) limited solutions for clinicians to prioritize the microenvironment of their own interests. Here we introduce SpatialFinder, a framework that combines a biomedical vision-language model (VLM) with a human-in-the-loop optimization pipeline to predict gene expression heterogeneity and rank high-value ROIs across routine H&E tissue slides. Evaluated across four Visium HD tissue types, SpatialFinder consistently outperforms VLM-only baselines for both diversity- and tumor-targeted ROI ranking, achieving Spearman's up to 0.89 and Overlap@10% up to 78.8%, an absolute 24.9 percentage-point gain over the strongest VLM. These results demonstrate the potential of human-AI collaboration to make spatial transcriptomics more cost-effective and clinically actionable.
Related Concept Videos
Improving Translational Accuracy
Improving Translational Accuracy
Ribosome Profiling
Applications of ribosome profiling
Ribosome profiling has many applications, including in vivo monitoring of translation inside a particular organ or tissue type and quantifying new protein synthesis levels.
The technique...

