RANet: Ranking Attention Network for Fast Video Object Segmentation
Ziqin Wang, Jun Xu, Li Liu, Fan Zhu, Ling Shao
Abstract
Despite online learning (OL) techniques have boosted the performance of semi-supervised video object segmentation (VOS) methods, the huge time costs of OL greatly restrict their practicality. Matching based and propagation based methods run at a faster speed by avoiding OL techniques. However, they are limited by sub-optimal accuracy, due to mismatching and drifting problems. In this paper, we develop a real-time yet very accurate Ranking Attention Network (RANet) for VOS. Specifically, to integrate the insights of matching based and propagation based methods, we employ an encoder-decoder framework to learn pixellevel similarity and segmentation in an end-to-end manner. To better utilize the similarity maps, we propose a novel ranking attention module, which automatically ranks and selects these maps for fine-grained VOS performance. Experiments on DAVIS 16 and DAVIS 17 datasets show that our RANet achieves the best speed-accuracy trade-off, e.g., with 33 milliseconds per frame and J &F=85.5% on DAVIS 16 . With OL, our RANet reaches J &F=87.1% on DAVIS 16 , exceeding state-of-the-art VOS methods. The code can be found at https://github.com/Storife/RANet .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers45
- EGNet: Edge Guidance Network for Salient Object DetectionJiaxing Zhao, Jiang-Jiang Liu, Deng-Ping Fan, Yang Cao et al.ICCV 2019 · 1,054 citations
- Rethinking Space-Time Networks with Improved Memory Coverage for Efficient Video Object SegmentationHo Kei Cheng, Yu-Wing Tai, Chi-Keung TangNeurIPS 2021 · 403 citations
- Zero-Shot Video Object Segmentation via Attentive Graph Neural NetworksWenguan Wang, Xiankai Lu, Jianbing Shen, David J. Crandall et al.ICCV 2019 · 294 citations
- MOSE: A New Dataset for Video Object Segmentation in Complex ScenesHenghui Ding, Chang Liu, Shuting He, Xudong Jiang et al.ICCV 2023 · 267 citations
- Video Object Segmentation with Adaptive Feature Bank and Uncertain-Region RefinementYongqing Liang, Xin Li, Navid H. Jafari, Jim ChenNeurIPS 2020 · 192 citations
Builds on3
- EGNet: Edge Guidance Network for Salient Object DetectionJiaxing Zhao, Jiang-Jiang Liu, Deng-Ping Fan, Yang Cao et al.ICCV 2019 · 1,054 citations
- Zero-Shot Video Object Segmentation via Attentive Graph Neural NetworksWenguan Wang, Xiankai Lu, Jianbing Shen, David J. Crandall et al.ICCV 2019 · 294 citations
- Towards Bridging Semantic Gap to Improve Semantic SegmentationYanwei Pang, Yazhao Li, Jianbing Shen, Ling ShaoICCV 2019 · 127 citations
Related papers
- DMVOS: Discriminative Matching for Real-time Video Object SegmentationPeisong Wen, Ruolin Yang, Qianqian Xu, Chen Qian et al.ACM MM 2020 · 17 citations
- Per-Clip Video Object SegmentationKwanyong Park, Sanghyun Woo, Seoung Wug Oh, In So Kweon et al.CVPR 2022 · 45 citations
- Boosting Video Object Segmentation via Space-Time Correspondence LearningYurong Zhang, Liulei Li, Wenguan Wang, Rong Xie et al.CVPR 2023
- Fast Video Object Segmentation With Temporal Aggregation Network and Dynamic Template MatchingXuhua Huang, Jiarui Xu, Yu-Wing Tai, Chi-Keung TangCVPR 2020
- SwiftNet: Real-Time Video Object SegmentationHaochen Wang, Xiaolong Jiang, Haibing Ren, Yao Hu et al.CVPR 2021
