Taming vision transformers for clinical laryngoscopy assessment

Xinzhu Zhang1, Jing Zhao1, Daoming Zong1

  • 1School of Computer Science and Technology, East China Normal University, North Zhongshan Road 3663, Shanghai, 200062, China.

PubMed
Summary

MedFormer, a Vision Transformer model, significantly improves early detection of laryngeal cancer (LCA) and precancerous lesions by analyzing laryngoscopic images. It outperforms other models and matches physician evaluations, aiding clinical diagnosis.