MoDAR: Using Motion Forecasting for 3D Object Detection in Point Cloud Sequences
Yingwei Li, Charles R. Qi, Yin Zhou, Chenxi Liu, Dragomir Anguelov
摘要
Occluded and long-range objects are ubiquitous and challenging for 3D object detection. Point cloud sequence data provide unique opportunities to improve such cases, as an occluded or distant object can be observed from different viewpoints or gets better visibility over time. However, the efficiency and effectiveness in encoding long-term sequence data can still be improved. In this work, we propose MoDAR, using motion forecasting outputs as a type of virtual modality, to augment LiDAR point clouds. The MoDAR modality propagates object information from temporal contexts to a target frame, represented as a set of virtual points, one for each object from a waypoint on a forecasted trajectory. A fused point cloud of both raw sensor points and the virtual points can then be fed to any off-the-shelf pointcloud based 3D object detector. Evaluated on the Waymo Open Dataset, our method significantly improves prior art detectors by using motion forecasting from extra-long sequences (e.g. 18 seconds), achieving new state of the arts, while not adding much computation overhead.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- T4P: Test-Time Training of Trajectory Prediction via Masked Autoencoder and Actor-Specific Token MemoryDaehee Park, Jaeseok Jeong, Sung-Hoon Yoon, Jaewoo Jeong 等CVPR 2024 · 被引用 14 次
- Towards Flexible 3D Perception: Object-Centric Occupancy Completion Augments 3D Object DetectionChaoda Zheng, Feng Wang, Naiyan Wang, Shuguang Cui 等NeurIPS 2024 · 被引用 5 次
- ForeSight: Multi-View Streaming Joint Object Detection and Trajectory ForecastingSandro Papais, Letian Wang, Brian Cheong, Steven L. WaslanderICCV 2025 · 被引用 1 次
- Scene Reconstruction as Mapping Priors for 3D DetectionYang Fu, Yuliang Zou, Hao Xiang, Xin Huang 等CVPR 2026 · 被引用 1 次
- Leveraging Temporal Cues for Semi-Supervised Multi-View 3D Object DetectionJinhyung Park, Navyata Sanghvi, Hiroki Adachi, Yoshihisa Shibata 等CVPR 2025
它引用的顶会 Paper26
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- Large Scale Interactive Motion Forecasting for Autonomous Driving : The Waymo Open Motion DatasetScott Ettinger, Shuyang Cheng, Benjamin Caine, Chenxi Liu 等ICCV 2021 · 被引用 817 次
- DenseTNT: End-to-end Trajectory Prediction from Dense Goal SetsJunru Gu, Chen Sun, Hang ZhaoICCV 2021 · 被引用 563 次
- DeepFusion: Lidar-Camera Deep Fusion for Multi-Modal 3D Object DetectionYingwei Li, Adams Wei Yu, Tianjian Meng, Benjamin Caine 等CVPR 2022 · 被引用 508 次
- Multimodal Virtual Point 3D DetectionTianwei Yin, Xingyi Zhou, Philipp KrähenbühlNeurIPS 2021 · 被引用 379 次
相关 Paper
- MGTANet: Encoding Sequential LiDAR Points Using Long Short-Term Motion-Guided Temporal Attention for 3D Object DetectionJunho Koh, Junhyung Lee, Youngwoo Lee, Jaekyum Kim 等AAAI 2023 · 被引用 34 次
- PointAugmenting: Cross-Modal Augmentation for 3D Object DetectionChunwei Wang, Chao Ma, Ming Zhu, Xiaokang YangCVPR 2021
- TREND: Unsupervised 3D Representation Learning via Temporal Forecasting for LiDAR PerceptionRunjian Chen, Hyoungseob Park, Bo Zhang, Wenqi Shao 等NeurIPS 2025 · 被引用 4 次
- PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object DetectionKuan-Chih Huang, Weijie Lyu, Ming-Hsuan Yang, Yi-Hsuan TsaiCVPR 2024
- 4D-Net for Learned Multi-Modal AlignmentA. J. Piergiovanni, Vincent Casser, Michael S. Ryoo, Anelia AngelovaICCV 2021 · 被引用 69 次
