Related Experiment Video
Updated: Aug 3, 2025

Estimation of Contact Regions Between Hands and Objects During Human Multi-Digit Grasping
Published on: April 21, 2023
Generalized Pose Decoupled Network for Unsupervised 3D Skeleton Sequence-Based Action Representation Learning
Mengyuan Liu1, Fanyang Meng2, Yongsheng Liang3
1Key Laboratory of Machine Perception, Peking University, Shenzhen Graduate School, Shenzhen, China.
Abstract:
Human action representation is derived from the description of human shape and motion. The traditional unsupervised 3-dimensional (3D) human action representation learning method uses a recurrent neural network (RNN)-based autoencoder to reconstruct the input pose sequence and then takes the midlevel feature of the autoencoder as representation. Although RNN can implicitly learn a certain amount of motion information, the extracted representation mainly describes the human shape and is insufficient to describe motion information. Therefore, we first present a handcrafted motion feature called pose flow to guide the reconstruction of the autoencoder, whose midlevel feature is expected to describe motion information. The performance is limited as we observe that actions can be distinctive in either motion direction or motion norm. For example, we can distinguish "sitting down" and "standing up" from motion direction yet distinguish "running" and "jogging" from motion norm. In these cases, it is difficult to learn distinctive features from pose flow where direction and norm are mixed. To this end, we present an explicit pose decoupled flow network (PDF-E) to learn from direction and norm in a multi-task learning framework, where 1 encoder is used to generate representation and 2 decoders are used to generating direction and norm, respectively. Further, we use reconstructing the input pose sequence as an additional constraint and present a generalized PDF network (PDF-G) to learn both motion and shape information, which achieves state-of-the-art performances on large-scale and challenging 3D action recognition datasets including the NTU RGB+D 60 dataset and NTU RGB+D 120 dataset.
Related Concept Videos
Muscle Coordination and Action
Agonists
Agonist muscles, often called prime movers, are the primary muscles responsible for producing a specific movement....
Generation of Action Potential in Skeletal Muscles
Like neurons, muscle cells are also regarded as excitable due to their capacity to change in response to stimuli, primarily due to voltage-gated ion channels embedded in their plasma membranes, which get activated by alterations in the...
Sequence Networks of Rotating Machines
Zero-sequence current induces a voltage drop across the generator's neutral impedance and other...
Carbon Skeletons
Structural Classification of Joints
A fibrous joint is where the adjacent bones are united by fibrous connective...
Relaxation of Skeletal Muscles
When an action potential reaches the axon terminal, it depolarizes the membrane and opens voltage-gated sodium channels. Sodium ions enter the cell, further depolarizing the presynaptic membrane. This depolarization causes voltage-gated calcium channels to open....

