TrajectoryFormer: 3D Object Tracking Transformer with Predictive Trajectory Hypotheses
Xuesong Chen, Shaoshuai Shi, Chao Zhang, Benjin Zhu, Qiang Wang, Ka Chun Cheung, Simon See, Hongsheng Li
摘要
3D multi-object tracking (MOT) is vital for many applications including autonomous driving vehicles and service robots. With the commonly used tracking-by-detection paradigm, 3D MOT has made important progress in recent years. However, these methods only use the detection boxes of the current frame to obtain trajectory-box association results, which makes it impossible for the tracker to recover objects missed by the detector. In this paper, we present Tra-jectoryFormer, a novel point-cloud-based 3D MOT framework. To recover the missed object by detector, we generates multiple trajectory hypotheses with hybrid candidate boxes, including temporally predicted boxes and currentframe detection boxes, for trajectory-box association. The predicted boxes can propagate object's history trajectory information to the current frame and thus the network can tolerate short-term miss detection of the tracked objects. We combine long-term object motion feature and short-term object appearance feature to create per-hypothesis feature embedding, which reduces the computational overhead for spatial-temporal encoding. Additionally, we introduce a Global-Local Interaction Module to conduct information interaction among all hypotheses and models their spatial relations, leading to accurate estimation of hypotheses. Our TrajectoryFormer achieves state-of-the-art performance on the Waymo 3D MOT benchmarks. Code is available at https://github.com/poodarchu/EFG .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Look Inside for More: Internal Spatial Modality Perception for 3D Anomaly DetectionHanzhe Liang, Guoyang Xie, Chengbin Hou, Bingshu Wang 等AAAI 2025 · 被引用 28 次
- M3Net: Multimodal Multi-task Learning for 3D Detection, Segmentation, and Occupancy Prediction in Autonomous DrivingXuesong Chen, Shaoshuai Shi, Tao Ma, Jingqiu Zhou 等AAAI 2025 · 被引用 14 次
- T4P: Test-Time Training of Trajectory Prediction via Masked Autoencoder and Actor-Specific Token MemoryDaehee Park, Jaeseok Jeong, Sung-Hoon Yoon, Jaewoo Jeong 等CVPR 2024 · 被引用 14 次
- ZOPP: A Framework of Zero-shot Offboard Panoptic Perception for Autonomous DrivingTao Ma, Hongbin Zhou, Qiusheng Huang, Xuemeng Yang 等NeurIPS 2024 · 被引用 8 次
- GSLAMOT: A Tracklet and Query Graph-based Simultaneous Locating, Mapping, and Multiple Object Tracking SystemShuo Wang, Yongcai Wang, Zhimin Xu, Yongyu Guo 等ACM MM 2024 · 被引用 6 次
它引用的顶会 Paper10
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen 等ICCV 2019 · 被引用 840 次
- Motion Transformer with Global Intention Localization and Local Movement RefinementShaoshuai Shi, Li Jiang, Dengxin Dai, Bernt SchieleNeurIPS 2022 · 被引用 515 次
- Quo Vadis: Is Trajectory Forecasting the Key Towards Long-Term Multi-Object Tracking?Patrick Dendorfer, Vladimir Yugay, Aljosa Osep, Laura Leal-TaixéNeurIPS 2022 · 被引用 77 次
- Forecasting from LiDAR via Future Object DetectionNeehar Peri, Jonathon Luiten, Mengtian Li, Aljosa Osep 等CVPR 2022 · 被引用 33 次
相关 Paper
- 3DMOTFormer: Graph Transformer for Online 3D Multi-Object TrackingShuxiao Ding, Eike Rehder, Lukas Schneider, Marius Cordts 等ICCV 2023 · 被引用 36 次
- PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object DetectionKuan-Chih Huang, Weijie Lyu, Ming-Hsuan Yang, Yi-Hsuan TsaiCVPR 2024
- DetZero: Rethinking Offboard 3D Object Detection with Long-term Sequential Point CloudsTao Ma, Xuemeng Yang, Hongbin Zhou, Xin Li 等ICCV 2023 · 被引用 46 次
- Delving into Motion-Aware Matching for Monocular 3D Object TrackingKuan-Chih Huang, Ming-Hsuan Yang, Yi-Hsuan TsaiICCV 2023 · 被引用 20 次
- TrackFormer: Multi-Object Tracking with TransformersTim Meinhardt, Alexander Kirillov, Laura Leal-Taixé, Christoph FeichtenhoferCVPR 2022 · 被引用 927 次
