PTTR: Relational 3D Point Cloud Object Tracking with Transformer
Changqing Zhou, Zhipeng Luo, Yueru Luo, Tianrui Liu, Liang Pan, Zhongang Cai, Haiyu Zhao, Shijian Lu
Abstract
In a point cloud sequence, 3D object tracking aims to predict the location and orientation of an object in the current search point cloud given a template point cloud. Motivated by the success of transformers, we propose Point Tracking TRansformer (PTTR), which efficiently predicts high-quality 3D tracking results in a coarse-to-fine manner with the help of transformer operations. PTTR consists of three novel designs. 1) Instead of random sampling, we design Relation-Aware Sampling to preserve relevant points to given templates during subsampling. 2) Furthermore, we propose a Point Relation Transformer (PRT) consisting of a self-attention and a cross-attention module. The global self-attention operation captures long-range dependencies to enhance encoded point features for the search area and the template, respectively. Subsequently, we generate the coarse tracking results by matching the two sets of point features via cross-attention. 3) Based on the coarse tracking results, we employ a novel Prediction Refinement Module to obtain the final refined prediction. In addition, we create a large-scale point cloud single object tracking benchmark based on the Waymo Open Dataset. Extensive experiments show that PTTR achieves superior point cloud tracking in both accuracy and efficiency. Our code is available at https://github.com/Jasonkks/PTTR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9bc7c7c6-1ca7-4bde-91ab-4c7e7f422c0cCited by top-tier papers26
- Accelerating DETR Convergence via Semantic-Aligned MatchingGongjie Zhang, Zhipeng Luo, Yingchen Yu, Kaiwen Cui et al.CVPR 2022 · 116 citations
- Cross-modal Orthogonal High-rank Augmentation for RGB-Event Transformer-trackersZhiyu Zhu, Junhui Hou, Dapeng Oliver WuICCV 2023 · 66 citations
- GLT-T: Global-Local Transformer Voting for 3D Single Object Tracking in Point CloudsJiahao Nie, Zhiwei He, Yuxiang Yang, Mingyu Gao et al.AAAI 2023 · 60 citations
- Learning Graph-embedded Key-event Back-tracing for Object Tracking in Event CloudsZhiyu Zhu, Junhui Hou, Xianqiang LyuNeurIPS 2022 · 48 citations
- MBPTrack: Improving 3D Point Cloud Tracking with Memory networks and Box PriorsTian-Xing Xu, Yuan-Chen Guo, Yu-Kun Lai, Song-Hai ZhangICCV 2023 · 34 citations
Builds on19
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
Related papers
- 3D Object Detection With PointformerXuran Pan, Zhuofan Xia, Shiji Song, Li Erran Li et al.CVPR 2021
- PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object DetectionKuan-Chih Huang, Weijie Lyu, Ming-Hsuan Yang, Yi-Hsuan TsaiCVPR 2024
- A Novel Object Re-Track Framework for 3D Point CloudsTuo Feng, Licheng Jiao, Hao Zhu, Long SunACM MM 2020 · 22 citations
- PointCFormer: A Relation-Based Progressive Feature Extraction Network for Point Cloud CompletionYi Zhong, Weize Quan, Dong-Ming Yan, Jie Jiang et al.AAAI 2025 · 3 citations
- Point Transformer V3: Simpler, Faster, StrongerXiaoyang Wu, Li Jiang, Peng-Shuai Wang, Zhijian Liu et al.CVPR 2024
