Alignment Before Aggregation: Trajectory Memory Retrieval Network for Video Object Segmentation
Rui Sun, Yuan Wang, Huayu Mai, Tianzhu Zhang, Feng Wu
摘要
Memory-based methods in semi-supervised video object segmentation task achieve competitive performance by performing dense matching between query and memory frames. However, most of the existing methods neglect the fact that videos carry rich temporal information yet redundant spatial information. In this case, direct pixel-level global matching will lead to ambiguous correspondences. In this work, we reconcile the inherent tension of spatial and temporal information to retrieve memory frame information along the object trajectory, and propose a novel and coherent Trajectory Memory Retrieval Network (TMRN) to equip with the trajectory information, including a spatial alignment module and a temporal aggregation module. The proposed TMRN enjoys several merits. First, TMRN is empowered to characterize the temporal correspondence which is in line with the nature of video in a data-driven manner. Second, we elegantly customize the spatial alignment module by coupling SVD initialization with agent-level correlation for representative agent construction and rectifying false matches caused by direct pairwise pixel-level correlation, respectively. Extensive experimental results on challenging benchmarks including DAVIS 2017 validation / test and Youtube-VOS 2018 / 2019 demonstrate that our TMRN, as a general plugin module, achieves consistent improvements over several leading methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- RankMatch: Exploring the Better Consistency Regularization for Semi-Supervised Semantic SegmentationHuayu Mai, Rui Sun, Tianzhu Zhang, Feng WuCVPR 2024 · 被引用 48 次
- DAW: Exploring the Better Weighting Function for Semi-supervised Semantic SegmentationRui Sun, Huayu Mai, Tianzhu Zhang, Feng WuNeurIPS 2023 · 被引用 40 次
- Focus on Query: Adversarial Mining Transformer for Few-Shot SegmentationYuan Wang, Naisong Luo, Tianzhu ZhangNeurIPS 2023 · 被引用 29 次
- Image-to-Image Matching via Foundation Models: A New Perspective for Open-Vocabulary Semantic SegmentationYuan Wang, Rui Sun, Naisong Luo, Yuwen Pan 等CVPR 2024 · 被引用 13 次
- Pay Attention to Target: Relation-Aware Temporal Consistency for Domain Adaptive Video Semantic SegmentationHuayu Mai, Rui Sun, Yuan Wang, Tianzhu Zhang 等AAAI 2024 · 被引用 12 次
它引用的顶会 Paper31
- Video Object Segmentation Using Space-Time Memory NetworksSeoung Wug Oh, Joon-Young Lee, Ning Xu, Seon Joo KimICCV 2019 · 被引用 845 次
- Rethinking Space-Time Networks with Improved Memory Coverage for Efficient Video Object SegmentationHo Kei Cheng, Yu-Wing Tai, Chi-Keung TangNeurIPS 2021 · 被引用 403 次
- Associating Objects with Transformers for Video Object SegmentationZongxin Yang, Yunchao Wei, Yi YangNeurIPS 2021 · 被引用 398 次
- Decoupling Features in Hierarchical Propagation for Video Object SegmentationZongxin Yang, Yi YangNeurIPS 2022 · 被引用 243 次
- Towards High-Resolution Salient Object DetectionYi Zeng, Pingping Zhang, Zhe Lin, Jianming Zhang 等ICCV 2019 · 被引用 232 次
相关 Paper
- Video Object Segmentation with Dynamic Memory Networks and Adaptive Object AlignmentShuxian Liang, Xu Shen, Jianqiang Huang, Xian-Sheng HuaICCV 2021 · 被引用 28 次
- Hierarchical Memory Matching Network for Video Object SegmentationHongje Seong, Seoung Wug Oh, Joon-Young Lee, Seongwon Lee 等ICCV 2021 · 被引用 126 次
- Dual Temporal Memory Network for Efficient Video Object SegmentationKaihua Zhang, Long Wang, Dong Liu, Bo Liu 等ACM MM 2020 · 被引用 16 次
- SWEM: Towards Real-Time Video Object Segmentation with Sequential Weighted Expectation-MaximizationZhihui Lin, Tianyu Yang, Maomao Li, Ziyu Wang 等CVPR 2022 · 被引用 42 次
- Efficient Regional Memory Network for Video Object SegmentationHaozhe Xie, Hongxun Yao, Shangchen Zhou, Shengping Zhang 等CVPR 2021
