Attention-Based Temporal Encoding Network with Background-Independent Motion Mask for Action Recognition

Zhengkui Weng1, Zhipeng Jin1, Shuangxi Chen1

  • 1Jiaxing Vocational and Technical College, Jiaxing, Zhejiang, China.

Summary

This study introduces an attention-based temporal encoding network (ATEN) with a background-independent motion mask (BIMM) for advanced video action recognition. The novel framework achieves high accuracy by focusing on critical motion segments and suppressing background noise.