Related Experiment Video

Updated: Sep 26, 2025

A Protocol for Real-time 3D Single Particle Tracking
10:16

A Protocol for Real-time 3D Single Particle Tracking

Published on: January 3, 2018

15.0K

Monocular Quasi-Dense 3D Object Tracking

Hou-Ning Hu, Yung-Hsu Yang, Tobias Fischer

    IEEE Transactions on Pattern Analysis and Machine Intelligence
    |April 19, 2022
    PubMed

    Abstract:

    A reliable and accurate 3D tracking framework is essential for predicting future locations of surrounding objects and planning the observer's actions in numerous applications such as autonomous driving. We propose a framework that can effectively associate moving objects over time and estimate their full 3D bounding box information from a sequence of 2D images captured on a moving platform. The object association leverages quasi-dense similarity learning to identify objects in various poses and viewpoints with appearance cues only. After initial 2D association, we further utilize 3D bounding boxes depth-ordering heuristics for robust instance association and motion-based 3D trajectory prediction for re-identification of occluded vehicles. In the end, an LSTM-based object velocity learning module aggregates the long-term trajectory information for more accurate motion extrapolation. Experiments on our proposed simulation data and real-world benchmarks, including KITTI, nuScenes, and Waymo datasets, show that our tracking framework offers robust object association and tracking on urban-driving scenarios. On the Waymo Open benchmark, we establish the first camera-only baseline in the 3D tracking and 3D detection challenges. Our quasi-dense 3D tracking pipeline achieves impressive improvements on the nuScenes 3D tracking benchmark with near five times tracking accuracy of the best vision-only submission among all published methods.

    More Related Videos

    3D Orbital Tracking in a Modified Two-photon Microscope: An Application to the Tracking of Intracellular Vesicles
    11:28

    3D Orbital Tracking in a Modified Two-photon Microscope: An Application to the Tracking of Intracellular Vesicles

    Published on: October 1, 2014

    10.3K
    Assessing Binocular Central Visual Field and Binocular Eye Movements in a Dichoptic Viewing Condition
    07:45

    Assessing Binocular Central Visual Field and Binocular Eye Movements in a Dichoptic Viewing Condition

    Published on: July 21, 2020

    4.6K

    Related Experiment Videos

    Last Updated: Sep 26, 2025

    A Protocol for Real-time 3D Single Particle Tracking
    10:16

    A Protocol for Real-time 3D Single Particle Tracking

    Published on: January 3, 2018

    15.0K
    3D Orbital Tracking in a Modified Two-photon Microscope: An Application to the Tracking of Intracellular Vesicles
    11:28

    3D Orbital Tracking in a Modified Two-photon Microscope: An Application to the Tracking of Intracellular Vesicles

    Published on: October 1, 2014

    10.3K
    Assessing Binocular Central Visual Field and Binocular Eye Movements in a Dichoptic Viewing Condition
    07:45

    Assessing Binocular Central Visual Field and Binocular Eye Movements in a Dichoptic Viewing Condition

    Published on: July 21, 2020

    4.6K

    Related Concept Videos

    Depth Perception and Spatial Vision01:15

    Depth Perception and Spatial Vision

    1.0K
    Depth perception is the ability to perceive objects three-dimensionally. It relies on two types of cues: binocular and monocular. Binocular cues depend on the combination of images from both eyes and how the eyes work together. Since the eyes are in slightly different positions, each eye captures a slightly different image. This disparity between images, known as binocular disparity, helps the brain interpret depth. When the brain compares these images, it determines the distance to an object.
    1.0K
    Curvilinear Motion: Rectangular Components01:23

    Curvilinear Motion: Rectangular Components

    694
    Curvilinear motion characterizes the movement of a particle or object along a curved path, notably evident when envisioning a car navigating a winding road. If the car starts at point A, its position vector is established within a fixed frame of reference, where the ratio of the position vector to its magnitude signifies the unit vector pointing in the position vector's direction.
    As the car advances, its position evolves over time. Quantifying the car's velocity involves computing the...
    694

    Articles linked to this work by shared authors, journal, and citation graph.

    Condition-Invariant Semantic Segmentation.

    IEEE transactions on pattern analysis and machine intelligence·2025

    PatCID: an open-access dataset of chemical structures in patent documents.

    Nature communications·2024

    QDTrack: Quasi-Dense Similarity Learning for Appearance-Only Multiple Object Tracking.

    IEEE transactions on pattern analysis and machine intelligence·2023

    Unifying Flow, Stereo and Depth Estimation.

    IEEE transactions on pattern analysis and machine intelligence·2023

    An ionic-liquid-modified melamine-formaldehyde aerogel for in-tube solid-phase microextraction of estrogens followed by high performance liquid chromatography with diode array detection.

    Mikrochimica acta·2019

    New Insights Into 3-Dimensional Anatomy of the Facial Mimetic Muscles Related to the Nasolabial Fold: An Iodine Staining Technique Based on Micro-Computed Tomography.

    Annals of plastic surgery·2019

    Enhanced Semantic Alignment in Transformer Tracking via Position Learning and Force-Directed Attention.

    IEEE transactions on pattern analysis and machine intelligence·2026

    Anticipating Object Interactions Via Aggregation and Distillation of Spatio-Temporal Knowledge From Vision Language Models.

    IEEE transactions on pattern analysis and machine intelligence·2026

    SPD Matrix Learning for Neuroimaging Analysis: Perspectives, Methods, and Challenges.

    IEEE transactions on pattern analysis and machine intelligence·2026

    Exploiting Vision Language Model for Training-Free 3D Point Cloud Understanding via Improved Graph Score Propagation.

    IEEE transactions on pattern analysis and machine intelligence·2026

    Flexible Motion Stylization via Multi-modality Latent Diffusion Model.

    IEEE transactions on pattern analysis and machine intelligence·2026

    Binarized High-Efficiency RAW Video Restoration and Beyond.

    IEEE transactions on pattern analysis and machine intelligence·2026

    Optimal navigation in two-dimensional flows: Control theory and reinforcement learning.

    Physical review. E·2026

    Cellular geometry compensates for Z-ring positioning in MapZ-deficient Streptococcus pneumoniae.

    The FEBS journal·2026

    A POST-PROCESSING METHOD FOR REFINEMENT OF CT-BASED DEEP LEARNING ACETABULAR SEGMENTATIONS.

    Osteoarthritis imaging·2026

    Structured navigation in a goal-directed task reveals flexible spatial coding.

    bioRxiv : the preprint server for biology·2026

    Research on Inclusive Wayfinding Systems in Transportation Hubs for Visually Impaired Older Adults.

    Journal of visualized experiments : JoVE·2026

    Hybrid frame-mask fixation for Gamma Knife radiosurgery of an extracranial vagal schwannoma: illustrative case.

    Journal of neurosurgery. Case lessons·2026
    See all related articles
    JoVE
    x logofacebook logolinkedin logoyoutube logo
    ABOUT JoVE
    OverviewLeadershipBlogJoVE Help Center
    AUTHORS
    Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
    LIBRARIANS
    TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
    RESEARCH
    JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
    EDUCATION
    JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
    Terms & Conditions of Use
    Privacy Policy
    Policies
    Jove
    Visualize
    Contact Us