Improving Laryngoscopy Image Analysis Through Integration of Global Information and Local Features in VoFoCD Dataset

Thao Thi Phuong Dao1,2,3,4, Tuan-Luc Huynh1,3, Minh-Khoi Pham5

  • 1University of Science, Ho Chi Minh City, Vietnam.

Summary

This study introduces a new dataset and a multitask AI model (MEAL) for analyzing laryngoscopy images. The model accurately detects vocal fold lesions and classifies images, aiding in diagnosing vocal fold disorders.