PointRePar : SpatioTemporal Point Relation Parsing for Robust Category-Unified 3D Tracking
Juntao Liu, Zikun Zhou, Zhuotao Tian, Guangming Lu, Jun Yu, Wenjie Pei
摘要
3D single object tracking (SOT) remains a highly challenging task due to the inherent crux in learning representations from point clouds to effectively capture both spatial shape features and temporal motion features. Most existing methods employ a category-specific optimization paradigm, training the tracking model individually for each object category to enhance tracking performance, albeit at the expense of generalizability across different categories. In this work, we propose a robust category-unified 3D SOT model, referred to as SpatioTemporal Point Relation Parsing model (PointRePar), which is capable of joint training across multiple categories while excelling in unified feature learning for both spatial shapes and temporal motions. Specifically, the proposed PointRePar captures and parses the latent point relations across both spatial and temporal domains to learn superior shape and motion characteristics for robust tracking. On the one hand, it models the multi-scale spatial point relations using a Mamba-based U-Net architecture with adaptive point-wise feature refinement. On the other hand, it captures both the point-level and box-level temporal relations to exploit the latent motion features. Extensive experiments across three benchmarks demonstrate that our PointRePar not only outperforms the existing category-unified 3D SOT methods significantly, but also compares favorably against the state-of-the-art category-specific methods. Codes will be released.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper17
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- PointMamba: A Simple State Space Model for Point Cloud AnalysisDingkang Liang, Xin Zhou, Wei Xu, Xingkui Zhu 等NeurIPS 2024 · 被引用 380 次
- Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space ModelYuheng Shi, Minjing Dong, Chang XuNeurIPS 2024 · 被引用 129 次
- PTTR: Relational 3D Point Cloud Object Tracking with TransformerChangqing Zhou, Zhipeng Luo, Yueru Luo, Tianrui Liu 等CVPR 2022 · 被引用 117 次
- Box-Aware Feature Enhancement for Single Object Tracking on Point CloudsChaoda Zheng, Xu Yan, Jiantao Gao, Weibing Zhao 等ICCV 2021 · 被引用 116 次
相关 Paper
- Generalizable Structure-Aware Keypoint Correspondence for Category-Unified 3D Single Object TrackingJie Xiao, Yinchao Ma, Yuyang Tang, Dengqing Yang 等CVPR 2026
- Towards Category Unification of 3D Single Object Tracking on Point CloudsJiahao Nie, Zhiwei He, Xudong Lv, Xueyi Zhou 等ICLR 2024 · 被引用 20 次
- TrackAny3D: Transferring Pretrained 3D Models for Category-Unified 3D Point Cloud TrackingMengmeng Wang, Haonan Wang, Yulong Li, Xiangjie Kong 等ICCV 2025 · 被引用 2 次
- M3SOT: Multi-Frame, Multi-Field, Multi-Space 3D Single Object TrackingJiaming Liu, Yue Wu, Maoguo Gong, Qiguang Miao 等AAAI 2024 · 被引用 17 次
- VoxelTrack: Exploring Multi-level Voxel Representation for 3D Point Cloud Object TrackingYuxuan Lu, Jiahao Nie, Zhiwei He, Hongjie Gu 等ACM MM 2024 · 被引用 4 次
