AVaTER: Fusing Audio, Visual, and Textual Modalities Using Cross-Modal Attention for Emotion Recognition

Avishek Das1, Moumita Sen Sarma1, Mohammed Moshiul Hoque1

  • 1Department of Computer Science and Engineering, Chittagong University of Engineering and Technology, Chittagong 4349, Bangladesh.

Sensors (Basel, Switzerland)
|September 28, 2024
PubMed