Accelerating SAM2 with Efficient Memory Attention Module via Spatiotemporal Token Pruning

Summary

A new method efficiently prunes memory tokens for SAM2, a video segmentation model. This accelerates inference by 1.7x with minimal performance impact, making video analysis faster.