Related Experiment Video
Updated: Apr 8, 2026

Combining Eye-tracking Data with an Analysis of Video Content from Free-viewing a Video of a Walk in an Urban Park Environment
Published on: May 7, 2019
Point Cloud Video Modeling With Progressive Prior Knowledge Guidance and Adaptive Neighboring Aggregation
None:
Point cloud video modeling not only has to address the natural irregularity of point clouds, but also the challenge of capturing spatial and temporal representation simultaneously. Current methods attempt to approximate the temporal dimension using several 3-D point cloud frame sequences but struggle in sparser conditions. Accurate point trajectory tracking is crucial for effectively capturing temporal dynamics, as point positions across different frames are often inconsistent, especially during rapid motion or at low frame rates. Conventional point tube operations aggregate motion features over fixed time windows but fail to capture rapidly changing scenes. Implicit tracking techniques are limited by quadratic time complexity, which restricts their practical use. In this article, we propose a native 4-D framework (N4DF) that guides the network to learn spatio-temporal dynamics from a native 4-D perspective. Furthermore, we devise a dynamic point spatio-temporal (DPST) convolution to adaptively select the optimal point-tracking strategy, which constructs local plane regions in anchor frames and propagates them to neighboring frames to evaluate point cross-frame movement distances. To further enhance the global modeling power of N4DF, we develop a dynamic self-tracking re-encoding (DSTR) module that employs point-wise self-attention to search for relevant points across the entire video. Compared with the recent 4-D modeling methods, N4DF demonstrates superior performance on MSR-Action3D and NTU RGB+D for action recognition (+0.7% and +1.2% accuracy, respectively), on HOI4D for action segmentation (+1% accuracy), and on Synthia 4-D and nuScenes-lidarseg for semantic segmentation (+0.49% and +1.7% mIoU, respectively). Our N4DF shows greater robustness at low frame-rate settings due to native 4-D modeling and adaptive tracking, making it suitable for tracking fast-moving objects in future real-time scenarios.
Related Concept Videos
Absolute Motion Analysis- General Plane Motion
As the drone's propellers rotate, an upward force is generated that counteracts the force of gravity, enabling the drone to lift off from the ground. This initial movement of the drone is along a straight path, representing a form of translational motion. In this phase, every point on the...
Depth Perception and Spatial Vision
Uniform Depth Channel Flow: Problem Solving
Relative Motion Analysis using Rotating Axes-Problem Solving
Here, in order to determine the magnitude of velocity and acceleration for point...
Relative Motion Analysis using Rotating Axes
However, to express the relative position of point B relative to point A, an additional frame of reference, denoted as x'y', is necessary. This additional frame not only translates but also rotates relative to the fixed frame, making it...
Observational Learning
