Related Experiment Video
Updated: Jul 5, 2026

11:34
High-resolution, High-speed, Three-dimensional Video Imaging with Digital Fringe Projection Techniques
Published on: December 3, 2013
Wavelet-based joint estimation and encoding of depth-image-based representations for free-viewpoint rendering
Matthieu Maitre1, Yoshihisa Shinagawa, Minh N Do
1Windows Experience Group, Microsoft, Redmont, WA 98052, USA. mmaitre@microsoft.com
Summary
This study introduces a new wavelet-based codec for depth-image-based rendering. It improves depth estimation and jointly encodes images and depth maps for efficient, scalable 3D viewing.
Area of Science:
- Computer Vision
- Image Processing
- Signal Processing
Background:
- Depth-image-based representation (DIBR) enables novel view synthesis.
- Accurate depth map estimation is crucial for realistic DIBR.
- Existing methods often struggle with depth ambiguity and efficient encoding.
Purpose of the Study:
- To develop a novel wavelet-based codec for static DIBR.
- To improve the joint estimation and encoding of depth maps.
- To achieve rate-distortion (RD) optimized performance and scalability.
Main Methods:
- A novel RD optimization scheme is proposed for joint depth estimation and encoding.
- A rate constraint is used to favor piecewise-smooth depth maps, reducing estimation ambiguity.
- Dynamic programming along trees of integer wavelet coefficients efficiently solves the optimization.
- Joint encoding of images and depth maps minimizes redundancy and optimizes bitrate allocation.
Main Results:
- The proposed codec effectively estimates and encodes depth maps from multiple views.
- The RD optimization scheme leads to improved depth map quality and efficient bitrate usage.
- The codec demonstrates scalability in both resolution and quality.
- Experiments on real data validate the effectiveness of the proposed approach.
Conclusions:
- The wavelet-based codec offers a significant advancement for static DIBR.
- Joint optimization of depth estimation and encoding enhances visual quality and compression efficiency.
- The codec's scalability makes it suitable for diverse applications requiring free viewpoint selection.
Related Concept Videos
Depth Perception and Spatial Vision
Depth perception is the ability to perceive objects three-dimensionally. It relies on two types of cues: binocular and monocular. Binocular cues depend on the combination of images from both eyes and how the eyes work together. Since the eyes are in slightly different positions, each eye captures a slightly different image. This disparity between images, known as binocular disparity, helps the brain interpret depth. When the brain compares these images, it determines the distance to an object.
Uniform Depth Channel Flow
Uniform depth channel flow keeps fluid depth consistent along channels such as irrigation canals. In natural channels, such as rivers, approximate uniform flow is often assumed. This condition occurs when the channel’s bottom slope matches the energy slope, balancing potential energy lost from gravity with head loss due to shear stress. This balance prevents depth changes along the channel length, resulting in a steady, uniform flow.Uniform flow in open channels with a constant cross-section...
Uniform Depth Channel Flow: Problem Solving
To calculate the flow rate for a trapezoidal channel, first, identify the bottom width, side slope, and flow depth of the channel. The cross-sectional area (A) corresponding to the depth of flow (y), channel bottom width (B), and side slope (θ) is determined by:Next, calculate the wetted perimeter, which includes the bottom width and the sloped side lengths in contact with the water. Using the values of the cross-sectional area and the wetted perimeter, determine the hydraulic radius by...
Deconvolution
Deconvolution, also known as inverse filtering, is the process of extracting the impulse response from known input and output signals. This technique is vital in scenarios where the system's characteristics are unknown, and they must be inferred from the observable signals.
Deconvolution involves several mathematical techniques to derive the impulse response. One common approach is polynomial division. In this method, the input and output sequences are treated as coefficients of...
Deconvolution involves several mathematical techniques to derive the impulse response. One common approach is polynomial division. In this method, the input and output sequences are treated as coefficients of...
Relative Motion Analysis using Rotating Axes
Consider a component AB undergoing a linear motion. Along with a linear motion, point B also rotates around point A. To comprehend this complex movement, position vectors for both points A and B are established using a stationary reference frame.
However, to express the relative position of point B relative to point A, an additional frame of reference, denoted as x'y', is necessary. This additional frame not only translates but also rotates relative to the fixed frame, making it instrumental in...
However, to express the relative position of point B relative to point A, an additional frame of reference, denoted as x'y', is necessary. This additional frame not only translates but also rotates relative to the fixed frame, making it instrumental in...
Divergence Theorem in 3D Space
In vector calculus, flux measures the total flow of a vector field through a surface. For a closed surface in three-dimensional space, this means measuring how much of the field passes outward through every point on the boundary. Directly calculating this flux can be difficult when the surface has a complicated or irregular shape. The Divergence Theorem provides a powerful alternative by relating surface flux to behavior inside the enclosed region.The Divergence Theorem states that the outward...