PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
Kuan-Chih Huang, Weijie Lyu, Ming-Hsuan Yang, Yi-Hsuan Tsai
Abstract
Recent temporal LiDAR-based 3D object detectors achieve promising performance based on the two-stage proposal-based approach. They generate 3D box candidates from the first-stage dense detector, followed by different temporal aggregation methods. However, these approaches require per-frame objects or whole point clouds, posing challenges related to memory bank utilization. Moreover, point clouds and trajectory features are combined solely based on concatenation, which may neglect effective interactions between them. In this paper, we propose a point-trajectory transformer with long short-term memory for efficient temporal 3D object detection. To this end, we only utilize point clouds of current-frame objects and their historical trajectories as input to minimize the memory bank storage requirement. Furthermore, we introduce modules to encode trajectory features, focusing on long short-term and future-aware perspectives, and then effectively aggregate them with point cloud features. We conduct extensive experiments on the large-scale Waymo dataset to demonstrate that our approach performs well against state-of-theart methods. Code and models will be made publicly available at https:// github.com/ kuanchihhuang/ PTT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- HGSFusion: Radar-Camera Fusion with Hybrid Generation and Synchronization for 3D Object DetectionZijian Gu, Jianwei Ma, Yan Huang, Honghao Wei et al.AAAI 2025 · 26 citations
- Towards a 3D Transfer-Based Black-Box Attack via Critical Feature GuidanceShuchao Pang, Zhenghan Chen, Shen Zhang, Liming Lu et al.ICCV 2025 · 1 citation
- Scene Reconstruction as Mapping Priors for 3D DetectionYang Fu, Yuliang Zou, Hao Xiang, Xin Huang et al.CVPR 2026 · 1 citation
- MAD: Memory-Augmented Detection of 3D ObjectsBen Agro, Sergio Casas, Patrick Wang, Thomas Gilles et al.CVPR 2025
- FASTer: Focal token Acquiring-and-Scaling Transformer for Long-term 3D Objection DetectionChenxu Dang, Zaipeng Duan, Pei An, Xinmin Zhang et al.CVPR 2025
Builds on22
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen et al.ICCV 2019 · 840 citations
- Voxel Transformer for 3D Object DetectionJiageng Mao, Yujing Xue, Minzhe Niu, Haoyue Bai et al.ICCV 2021 · 535 citations
- Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point CloudsYifan Zhang, Qingyong Hu, Guoquan Xu, Yanxin Ma et al.CVPR 2022 · 376 citations
- Improving 3D Object Detection with Channel-wise TransformerHualian Sheng, Sijia Cai, Yuan Liu, Bing Deng et al.ICCV 2021 · 293 citations
Related papers
- LiDAR-Based Online 3D Video Object Detection With Graph-Based Message Passing and Spatiotemporal Transformer AttentionJunbo Yin, Jianbing Shen, Chenye Guan, Dingfu Zhou et al.CVPR 2020
- TrajectoryFormer: 3D Object Tracking Transformer with Predictive Trajectory HypothesesXuesong Chen, Shaoshuai Shi, Chao Zhang, Benjin Zhu et al.ICCV 2023 · 25 citations
- PTTR: Relational 3D Point Cloud Object Tracking with TransformerChangqing Zhou, Zhipeng Luo, Yueru Luo, Tianrui Liu et al.CVPR 2022 · 117 citations
- PTNET: A Proposal-Centric Transformer Net-Work for 3D Object DetectionJianping Zhong, Zhaobo Qi, Kaiwen Duan, Xinyan Liu et al.ICLR 2026
- MoDAR: Using Motion Forecasting for 3D Object Detection in Point Cloud SequencesYingwei Li, Charles R. Qi, Yin Zhou, Chenxi Liu et al.CVPR 2023
