Local Dimension Enhancement Representation Learning for Skeleton-Based Action Segmentation

Summary

Self-supervised learning for skeleton-based temporal action segmentation struggles with short-term motion. The Local Dimension Enhancement (LoDE) framework improves this by introducing motion units and multi-scale learning to reduce local dimension collapse.