Related Experiment Video
Updated: Mar 29, 2026

Author Spotlight: Insights into Visual Cortex Research Through Wide-View fMRI Mapping
Published on: December 8, 2023
A novel superpixel based Vision Transformer for improving interpretability in glaucoma screening
Jorge Hernández1, Silvia Alayón2, José F Sigut2
1Department of Computer Science and Systems Engineering, University of La Laguna, 38200, San Cristóbal de La Laguna, Santa Cruz de Tenerife, Spain. jhernanv@ull.edu.es.
Abstract:
Interpretability remains one of the major challenges in the clinical adoption of deep learning models for medical image analysis. In ophthalmology, particularly for glaucoma screening, explainable artificial intelligence (XAI) methods are essential for ensuring trust and diagnostic transparency. This study introduces the Superpixel-based Vision Transformer (SpxViT), a model designed to enhance interpretability while maintaining competitive accuracy. SpxViT replaces the traditional fixed grid tokenization of Vision Transformers. (ViTs) with a superpixel-based approach that preserves semantic boundaries within the retinal image. Two variants, SpxViT_fix and SpxViT_var, were evaluated on public and private glaucoma datasets. Results demonstrate that SpxViT achieves comparable accuracy to ViT-B/16 (91.9% vs. 92.5%) while producing more clinically consistent attention maps focused on the optic disc and cup.
More Related Videos
07:11Assessing Early Stage Open-Angle Glaucoma in Patients by Isolated-Check Visual Evoked Potential
Published on: May 25, 2020
07:12Development of a Gaze-Contingent Display Framework Designed for Perceptual and Oculomotor Research with Simulated Central Vision Loss
Published on: April 11, 2025
Related Concept Videos
Glaucoma: Overview
Open Angle Glaucoma: Treatment
Drugs such as carbonic anhydrase inhibitors, α2- and...
Angle Closure Glaucoma: Treatment