Attention-Guided Disentangled Feature Aggregation for Video Object Detection

Shishir Muralidhara1,2, Khurram Azeem Hashmi1,2,3, Alain Pagani3

  • 1Department of Computer Science, Technical University of Kaiserslautern, 67663 Kaiserslautern, Germany.

Summary

This study introduces an attention-heavy framework for video object detection, improving accuracy by disentangling and aggregating frame features. The novel approach enhances object localization and classification in challenging video sequences.