Delving into Dynamic Scene Cue-Consistency for Robust 3D Multi-Object Tracking
Haonan Zhang, Xinyao Wang, Boxi Wu, Tu Zheng, Wang Yunhua, Zheng Yang
Abstract
3D multi-object tracking is a critical and challenging task in the field of autonomous driving. A common paradigm relies on modeling individual object motion, e.g., Kalman filters, to predict trajectories. While effective in simple scenarios, this approach often struggles in crowded environments or with inaccurate detections, as it overlooks the rich geometric relationships between objects. This highlights the need to leverage spatial cues. However, existing geometry-aware methods can be susceptible to interference from irrelevant objects, leading to ambiguous features and incorrect associations. To address this, we propose focusing on cue-consistency: identifying and matching stable spatial patterns over time. We introduce the Dynamic Scene Cue-Consistency Tracker (DSC-Track) to implement this principle. Firstly, we design a unified spatiotemporal encoder using Point Pair Features (PPF) to learn discriminative trajectory embeddings while suppressing interference. Secondly, our cue-consistency transformer module explicitly aligns consistent feature representations between historical tracks and current detections. Finally, a dynamic update mechanism preserves salient spatiotemporal information for stable online tracking. Extensive experiments on the nuScenes and Waymo Open Datasets validate the effectiveness and robustness of our approach. On the nuScenes benchmark, for instance, our method achieves state-of-theart performance, reaching 73.2% and 70.3% AMOTA on the validation and test sets, respectively. Code Repository: https://github.com/zhn12343333/DSCTrack
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on12
- Deep Closest Point: Learning Representations for Point Cloud RegistrationYue Wang, Justin SolomonICCV 2019 · 1,026 citations
- TrajectoryFormer: 3D Object Tracking Transformer with Predictive Trajectory HypothesesXuesong Chen, Shaoshuai Shi, Chao Zhang, Benjin Zhu et al.ICCV 2023 · 25 citations
- Delving into Motion-Aware Matching for Monocular 3D Object TrackingKuan-Chih Huang, Ming-Hsuan Yang, Yi-Hsuan TsaiICCV 2023 · 20 citations
- Rotation-Invariant Transformer for Point Cloud MatchingHao Yu, Zheng Qin, Ji Hou, Mahdi Saleh et al.CVPR 2023
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora et al.CVPR 2020
Related papers
- GRAE-3DMOT: Geometry Relation-Aware Encoder for Online 3D Multi-Object TrackingHyunseop Kim, Hyo-Jun Lee, Yonguk Lee, Jinu Lee et al.CVPR 2025
- Standing Between Past and Future: Spatio-Temporal Modeling for Multi-Camera 3D Multi-Object TrackingZiqi Pang, Jie Li, Pavel Tokmakov, Dian Chen et al.CVPR 2023
- S2-Track: A Simple yet Strong Approach for End-to-End 3D Multi-Object TrackingTao Tang, Lijun Zhou, Pengkun Hao, Zihang He et al.ICML 2025
- Time3D: End-to-End Joint Monocular 3D Object Detection and Tracking for Autonomous DrivingPeixuan Li, Jieyu JinCVPR 2022 · 52 citations
- CXTrack: Improving 3D Point Cloud Tracking with Contextual InformationTian-Xing Xu, Yuan-Chen Guo, Yu-Kun Lai, Song-Hai ZhangCVPR 2023
