Enhanced Hybrid Vision Transformer with Multi-Scale Feature Integration and Patch Dropping for Facial Expression

Nianfeng Li1, Yongyuan Huang1, Zhenyan Wang1

  • 1College of Computer Science and Technology, Changchun University, No. 6543, Satellite Road, Changchun 130022, China.

PubMed
Summary

This study introduces a lightweight hybrid vision transformer for facial expression recognition (FER). The novel method enhances feature extraction and reduces computational costs, achieving high accuracy on benchmark datasets.

Related Concept Videos