Related Experiment Video
Updated: Aug 25, 2025

11:34
High-resolution, High-speed, Three-dimensional Video Imaging with Digital Fringe Projection Techniques
Published on: December 3, 2013
15.7K
RobustFusion: Robust Volumetric Performance Reconstruction Under Human-Object Interactions From Monocular RGBD
IEEE Transactions on Pattern Analysis and Machine Intelligence
|October 19, 2022
Summary
RobustFusion reconstructs 4D human performance during object interactions using a single RGBD sensor. This system effectively handles complex interactions and occlusions for enhanced virtual and augmented reality applications.
Area of Science:
- Computer Vision
- Computer Graphics
- Human-Computer Interaction
Background:
- High-quality 4D reconstruction of human performance with object interactions is crucial for immersive virtual and augmented reality (VR/AR).
- Existing methods struggle with complex interactions and occlusions, particularly in monocular settings.
Purpose of the Study:
- To propose RobustFusion, a robust volumetric performance reconstruction system for human-object interaction using a single RGBD sensor.
- To address challenges of complex interactions and severe occlusions in monocular 4D reconstruction.
Main Methods:
- A semantic-aware scene decoupling scheme for explicit occlusion modeling.
- Segmentation refinement and robust object tracking for temporal consistency.
- A data-driven performance capture scheme with spatial relation priors and interaction cues.
- Adaptive fusion with occlusion analysis and human parsing for coherent reconstruction.
Main Results:
- RobustFusion achieves high-quality 4D human performance reconstruction in complex human-object interaction scenarios.
- The system effectively handles severe occlusions and disentanglement uncertainty.
- Maintains temporal consistency and natural motion, even in occluded regions.
Conclusions:
- RobustFusion provides a lightweight yet effective solution for 4D human performance reconstruction from a single RGBD sensor.
- The proposed methods significantly improve reconstruction quality under challenging human-object interaction conditions.
- Enables more realistic and immersive VR/AR experiences through accurate human-object interaction capture.
Related Concept Videos
Depth Perception and Spatial Vision
837
Depth perception is the ability to perceive objects three-dimensionally. It relies on two types of cues: binocular and monocular. Binocular cues depend on the combination of images from both eyes and how the eyes work together. Since the eyes are in slightly different positions, each eye captures a slightly different image. This disparity between images, known as binocular disparity, helps the brain interpret depth. When the brain compares these images, it determines the distance to an object.
837
Uniform Depth Channel Flow: Problem Solving
116
To calculate the flow rate for a trapezoidal channel, first, identify the bottom width, side slope, and flow depth of the channel. The cross-sectional area (A) corresponding to the depth of flow (y), channel bottom width (B), and side slope (θ) is determined by:Next, calculate the wetted perimeter, which includes the bottom width and the sloped side lengths in contact with the water. Using the values of the cross-sectional area and the wetted perimeter, determine the hydraulic radius by...
116
Uniform Depth Channel Flow
126
Uniform depth channel flow keeps fluid depth consistent along channels such as irrigation canals. In natural channels, such as rivers, approximate uniform flow is often assumed. This condition occurs when the channel’s bottom slope matches the energy slope, balancing potential energy lost from gravity with head loss due to shear stress. This balance prevents depth changes along the channel length, resulting in a steady, uniform flow.Uniform flow in open channels with a constant cross-section...
126

