Fast Template Matching and Update for Video Object Tracking and Segmentation
Mingjie Sun, Jimin Xiao, Eng Gee Lim, Bingfeng Zhang, Yao Zhao
摘要
In this paper, the main task we aim to tackle is the multiinstance semi-supervised video object segmentation across a sequence of frames where only the first-frame box-level ground-truth is provided. Detection-based algorithms are widely adopted to handle this task, and the challenges lie in the selection of the matching method to predict the result as well as to decide whether to update the target template using the newly predicted result. The existing methods, however, make these selections in a rough and inflexible way, compromising their performance. To overcome this limitation, we propose a novel approach which utilizes reinforcement learning to make these two decisions at the same time. Specifically, the reinforcement learning agent learns to decide whether to update the target template according to the quality of the predicted result. The choice of the matching method will be determined at the same time, based on the action history of the reinforcement learning agent. Experiments show that our method is almost 10 times faster than the previous state-of-the-art method with even higher accuracy (region similarity of 69.1% on DAVIS 2017 dataset).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- MOSE: A New Dataset for Video Object Segmentation in Complex ScenesHenghui Ding, Chang Liu, Shuting He, Xudong Jiang 等ICCV 2023 · 被引用 267 次
- Locality-Aware Inter-and Intra-Video Reconstruction for Self-Supervised Correspondence LearningLiulei Li, Tianfei Zhou, Wenguan Wang, Lu Yang 等CVPR 2022 · 被引用 41 次
- Video Object Segmentation with Dynamic Memory Networks and Adaptive Object AlignmentShuxian Liang, Xu Shen, Jianqiang Huang, Xian-Sheng HuaICCV 2021 · 被引用 28 次
- Point-VOS: Pointing Up Video Object SegmentationSabarinath Mahadevan, Idil Esen Zulfikar, Paul Voigtlaender, Bastian LeibeCVPR 2024 · 被引用 3 次
- RELO: Reinforcement Learning to Localize for Visual Object TrackingXin Chen, Chuanyu Sun, Jiao Xu, Houwen Peng 等ICML 2026 · 被引用 1 次
它引用的顶会 Paper2
相关 Paper
- Target-Aware Object Discovery and Association for Unsupervised Video Multi-Object SegmentationTianfei Zhou, Jianwu Li, Xueyi Li, Ling ShaoCVPR 2021
- A Transductive Approach for Video Object SegmentationYizhuo Zhang, Zhirong Wu, Houwen Peng, Stephen LinCVPR 2020
- Learning To Recommend Frame for Interactive Video Object Segmentation in the WildZhaoyuan Yin, Jia Zheng, Weixin Luo, Shenhan Qian 等CVPR 2021
- Joint Inductive and Transductive Learning for Video Object SegmentationYunyao Mao, Ning Wang, Wengang Zhou, Houqiang LiICCV 2021 · 被引用 111 次
- Learning Dynamic Network Using a Reuse Gate Function in Semi-Supervised Video Object SegmentationHyojin Park, Jayeon Yoo, Seohyeong Jeong, Ganesh Venkatesh 等CVPR 2021
