DepthTrack: Unveiling the Power of RGBD Tracking
Song Yan, Jinyu Yang, Jani Käpylä, Feng Zheng, Ales Leonardis, Joni-Kristian Kämäräinen
Abstract
RGBD (RGB plus depth) object tracking is gaining momentum as RGBD sensors have become popular in many application fields such as robotics. However, the best RGBD trackers are extensions of the state-of-the-art deep RGB trackers. They are trained with RGB data and the depth channel is used as a sidekick for subtleties such as occlusion detection. This can be explained by the fact that there are no sufficiently large RGBD datasets to 1) train "deep depth trackers" and to 2) challenge RGB trackers with sequences for which the depth cue is essential. This work introduces a new RGBD tracking dataset -Depth-Track -that has twice as many sequences (200) and scene types (40) than in the largest existing dataset, and three times more objects (90). In addition, the average length of the sequences (1473), the number of deformable objects ( 16 ) and the number of annotated tracking attributes (15) have been increased. Furthermore, by running the SotA RGB and RGBD trackers on DepthTrack, we propose a new RGBD tracking baseline, namely DeT, which reveals that deep RGBD tracking indeed benefits from genuine training data. The code and dataset is available at https://github.com/xiaozai/DeT .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 60fc517a-4640-4232-80dd-153e95c9ecf3Cited by top-tier papers28
- Prompting for Multi-Modal TrackingJinyu Yang, Zhe Li, Feng Zheng, Ales Leonardis et al.ACM MM 2022 · 167 citations
- DFormer: Rethinking RGBD Representation Learning for Semantic SegmentationBowen Yin, Xuying Zhang, Zhong-Yu Li, Li Liu et al.ICLR 2024 · 110 citations
- RGBD1K: A Large-Scale Dataset and Benchmark for RGB-D Object TrackingXuefeng Zhu, Tianyang Xu, Zhangyong Tang, Zucheng Wu et al.AAAI 2023 · 79 citations
- Single-Model and Any-Modality for Video Object TrackingZongwei Wu, Jilai Zheng, Xiangxuan Ren, Florin-Alexandru Vasluianu et al.CVPR 2024 · 78 citations
- Generative-Based Fusion Mechanism for Multi-Modal TrackingZhangyong Tang, Tianyang Xu, Xiaojun Wu, Xuefeng Zhu et al.AAAI 2024 · 78 citations
Builds on4
- Learning Discriminative Model Prediction for TrackingGoutam Bhat, Martin Danelljan, Luc Van Gool, Radu TimofteICCV 2019 · 1,294 citations
- CDTB: A Color and Depth Visual Object Tracking Dataset and BenchmarkAlan Lukezic, Ugur Kart, Jani Käpylä, Ahmed Durmush et al.ICCV 2019 · 79 citations
- Alpha-Refine: Boosting Tracking Performance by Precise Bounding Box EstimationBin Yan, Xinyu Zhang, Dong Wang, Huchuan Lu et al.CVPR 2021
- D3S - A Discriminative Single Shot Segmentation TrackerAlan Lukezic, Jiri Matas, Matej KristanCVPR 2020
Related papers
- Resource-Efficient RGBD Aerial TrackingJinyu Yang, Shang Gao, Zhe Li, Feng Zheng et al.CVPR 2023
- Event6D: Event-based Novel Object 6D Pose TrackingJae-Young Kang, Hoonhee Cho, Taeyeop Lee, Minjun Kang et al.CVPR 2026 · 4 citations
- GSOT3D: Towards Generic 3D Single Object Tracking in the WildYifan Jiao, Yunhao Li, Junhua Ding, Qing Yang et al.ICCV 2025
- Cross-Modal Object Tracking: Modality-Aware Representations and a Unified BenchmarkChenglong Li, Tianhao Zhu, Lei Liu, Xiaonan Si et al.AAAI 2022 · 12 citations
- Is Depth Really Necessary for Salient Object Detection?Jiawei Zhao, Yifan Zhao, Jia Li, Xiaowu ChenACM MM 2020 · 71 citations
