Computational framework for fusing eye movements and spoken narratives for image annotation

Preethi Vaidyanathan1, Emily Prud'hommeaux2, Cecilia O Alm3

  • 1Eyegaze Inc., Fairfax, VA, USA.

Journal of Vision
|July 18, 2020
PubMed
Summary

This study integrates human gaze and spoken language to automatically label important image regions. This approach enhances computer vision by bridging the gap between machine processing and human understanding of visual data.