Online Decision Based Visual Tracking via Reinforcement Learning
Ke Song, Wei Zhang, Ran Song, Yibin Li
Abstract
A deep visual tracker is typically based on either object detection or template matching while each of them is only suitable for a particular group of scenes. It is straightforward to consider fusing them together to pursue more reliable tracking. However, this is not wise as they follow different tracking principles. Unlike previous fusion-based methods, we propose a novel ensemble framework, named DTNet, with an online decision mechanism for visual tracking based on hierarchical reinforcement learning. The decision mechanism substantiates an intelligent switching strategy where the detection and the template trackers have to compete with each other to conduct tracking within different scenes that they are adept in. Besides, we present a novel detection tracker which avoids the common issue of incorrect proposal. Extensive results show that our DTNet achieves stateof-the-art tracking performance as well as a good balance between accuracy and efficiency. The project website is available at https://vsislab.github. io/DTNet/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Temporal Complementarity-Guided Reinforcement Learning for Image-to-Video Person Re-IdentificationWei Wu, Jiawei Liu, Kecheng Zheng, Qibin Sun et al.CVPR 2022 · 17 citations
- Open-World Drone Active Tracking with Goal-Centered RewardsHaowei Sun, Jinwu Hu, Zhirui Zhang, Haoyuan Tian et al.NeurIPS 2025 · 7 citations
- RELO: Reinforcement Learning to Localize for Visual Object TrackingXin Chen, Chuanyu Sun, Jiao Xu, Houwen Peng et al.ICML 2026 · 1 citation
Builds on1
Related papers
- POST: POlicy-Based Switch TrackingNing Wang, Wengang Zhou, Guojun Qi, Houqiang LiAAAI 2020 · 9 citations
- Visual Tracking via Hierarchical Deep Reinforcement LearningDawei Zhang, Zhonglong Zheng, Riheng Jia, Minglu LiAAAI 2021 · 29 citations
- Learning the Model Update for Siamese TrackersLichao Zhang, Abel Gonzalez-Garcia, Joost van de Weijer, Martin Danelljan et al.ICCV 2019 · 371 citations
- Fast Template Matching and Update for Video Object Tracking and SegmentationMingjie Sun, Jimin Xiao, Eng Gee Lim, Bingfeng Zhang et al.CVPR 2020
- DreamTrack: Dreaming the Future for Multimodal Visual Object TrackingMingzhe Guo, Weiping Tan, Wenyu Ran, Liping Jing et al.CVPR 2025
