Deformable Siamese Attention Networks for Visual Object Tracking
Yuechen Yu, Yilei Xiong, Weilin Huang, Matthew R. Scott
Abstract
Siamese-based trackers have achieved excellent performance on visual object tracking. However, the target template is not updated online, and the features of the target template and search image are computed independently in a Siamese architecture. In this paper, we propose Deformable Siamese Attention Networks, referred to as Sia-mAttn, by introducing a new Siamese attention mechanism that computes deformable self-attention and crossattention. The self-attention learns strong context information via spatial attention, and selectively emphasizes interdependent channel-wise features with channel attention. The cross-attention is capable of aggregating rich contextual interdependencies between the target template and the search image, providing an implicit manner to adaptively update the target template. In addition, we design a region refinement module that computes depth-wise cross correlations between the attentional features for more accurate tracking. We conduct experiments on six benchmarks, where our method achieves new state-of-the-art results, outperforming the strong baseline, SiamRPN++ [24], by 0.464→0.537 and 0.415→0.470 EAO on VOT 2016 and 2018. Our code is available at: https://github.com/msight- tech/research-siamattn.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 19313b3d-43ee-44bf-8220-03cb28c81edaCited by top-tier papers38
- Learning Spatio-Temporal Transformer for Visual TrackingBin Yan, Houwen Peng, Jianlong Fu, Dong Wang et al.ICCV 2021 · 1,062 citations
- MixFormer: End-to-End Tracking with Iterative Mixed AttentionYutao Cui, Cheng Jiang, Limin Wang, Gangshan WuCVPR 2022 · 746 citations
- SwinTrack: A Simple and Strong Baseline for Transformer TrackingLiting Lin, Heng Fan, Zhipeng Zhang, Yong Xu et al.NeurIPS 2022 · 556 citations
- Transforming Model Prediction for TrackingChristoph Mayer, Martin Danelljan, Goutam Bhat, Matthieu Paul et al.CVPR 2022 · 399 citations
- Learning Target Candidate Association to Keep Track of What Not to TrackChristoph Mayer, Martin Danelljan, Danda Pani Paudel, Luc Van GoolICCV 2021 · 356 citations
Builds on6
- Learning Discriminative Model Prediction for TrackingGoutam Bhat, Martin Danelljan, Luc Van Gool, Radu TimofteICCV 2019 · 1,294 citations
- Learning the Model Update for Siamese TrackersLichao Zhang, Abel Gonzalez-Garcia, Joost van de Weijer, Martin Danelljan et al.ICCV 2019 · 371 citations
- Learning Aberrance Repressed Correlation Filters for Real-Time UAV TrackingZiyuan Huang, Changhong Fu, Yiming Li, Fuling Lin et al.ICCV 2019 · 347 citations
- GradNet: Gradient-Guided Network for Visual Object TrackingPeixia Li, Boyu Chen, Wanli Ouyang, Dong Wang et al.ICCV 2019 · 255 citations
- Joint Group Feature Selection and Discriminative Filter Learning for Robust Visual Object TrackingTianyang Xu, Zhenhua Feng, Xiao-Jun Wu, Josef KittlerICCV 2019 · 182 citations
Related papers
- Discriminative and Robust Online Learning for Siamese Visual TrackingJinghao Zhou, Peng Wang, Haoyang SunAAAI 2020 · 66 citations
- Graph Attention TrackingDongyan Guo, Yanyan Shao, Ying Cui, Zhenhua Wang et al.CVPR 2021
- Correlation-Aware Deep TrackingFei Xie, Chunyu Wang, Guangting Wang, Yue Cao et al.CVPR 2022 · 189 citations
- Reinforced Similarity Learning: Siamese Relation Networks for Robust Object TrackingDawei Zhang, Zhonglong Zheng, Minglu Li, Xiaowei He et al.ACM MM 2020 · 15 citations
- STMTrack: Template-Free Visual Tracking With Space-Time Memory NetworksZhihong Fu, Qingjie Liu, Zehua Fu, Yunhong WangCVPR 2021
