Joint 3D Object Detection and Tracking Using Spatio-Temporal Representation of Camera Image and LiDAR Point Clouds
Junho Koh, Jaekyum Kim, Jin Hyeok Yoo, Yecheol Kim, Dongsuk Kum, Jun Won Choi
摘要
In this paper, we propose a new joint object detection and tracking (JoDT) framework for 3D object detection and tracking based on camera and LiDAR sensors. The proposed method, referred to as 3D DetecTrack, enables the detector and tracker to cooperate to generate a spatio-temporal representation of the camera and LiDAR data, with which 3D object detection and tracking are then performed. The detector constructs the spatio-temporal features via the weighted temporal aggregation of the spatial features obtained by the camera and LiDAR fusion. Then, the detector reconfigures the initial detection results using information from the tracklets maintained up to the previous time step. Based on the spatio-temporal features generated by the detector, the tracker associates the detected objects with previously tracked objects using a graph neural network (GNN). We devise a fully-connected GNN facilitated by a combination of rule-based edge pruning and attention-based edge gating, which exploits both spatial and temporal object contexts to improve tracking performance. The experiments conducted on both KITTI and nuScenes benchmarks demonstrate that the proposed 3D DetecTrack achieves significant improvements in both detection and tracking performances over baseline methods and achieves state-of-the-art performance among existing methods through collaboration between the detector and tracker.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Multi-Scene Generalized Trajectory Global Graph Solver with Composite Nodes for Multiple Object TrackingYan Gao, Haojun Xu, Jie Li, Nannan Wang 等AAAI 2024 · 被引用 17 次
- msLPCC: A Multimodal-Driven Scalable Framework for Deep LiDAR Point Cloud CompressionMiaohui Wang, Runnan Huang, Hengjin Dong, Di Lin 等AAAI 2024 · 被引用 7 次
- 3D Measurement of Complex Textured Objects Based on Bidirectional Fringe ProjectionYuchong Chen, Jian Yu, Shaoyan Gai, Zeyu Cai 等AAAI 2025 · 被引用 4 次
- High-Precision 3D Measurement of Complex Textured Surfaces Using Multiple Filtering ApproachYuchong Chen, Jian Yu, Shaoyan Gai, Zeyu Cai 等ICCV 2025
它引用的顶会 Paper7
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen 等ICCV 2019 · 被引用 840 次
- CIA-SSD: Confident IoU-Aware Single-Stage Object Detector From Point CloudWu Zheng, Weiliang Tang, Sijin Chen, Li Jiang 等AAAI 2021 · 被引用 335 次
- Joint Monocular 3D Vehicle Detection and TrackingHou-Ning Hu, Qi-Zhi Cai, Dequan Wang, Ji Lin 等ICCV 2019 · 被引用 242 次
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora 等CVPR 2020
- RetinaTrack: Online Single Stage Joint Detection and TrackingZhichao Lu, Vivek Rathod, Ronny Votel, Jonathan HuangCVPR 2020
相关 Paper
- GNN3DMOT: Graph Neural Network for 3D Multi-Object Tracking With 2D-3D Multi-Feature LearningXinshuo Weng, Yongxin Wang, Yunze Man, Kris M. KitaniCVPR 2020
- PC-RGNN: Point Cloud Completion and Graph Neural Network for 3D Object DetectionYanan Zhang, Di Huang, Yunhong WangAAAI 2021 · 被引用 109 次
- Point-GNN: Graph Neural Network for 3D Object Detection in a Point CloudWeijing Shi, Raj RajkumarCVPR 2020
- Exploring Simple 3D Multi-Object Tracking for Autonomous DrivingChenxu Luo, Xiaodong Yang, Alan L. YuilleICCV 2021 · 被引用 122 次
- GraphAlign: Enhancing Accurate Feature Alignment by Graph matching for Multi-Modal 3D Object DetectionZiying Song, Haiyue Wei, Lin Bai, Lei Yang 等ICCV 2023 · 被引用 73 次
