Memory-Augmented Non-Local Attention for Video Super-Resolution
Jiyang Yu, Jingen Liu, Liefeng Bo, Tao Mei
Abstract
In this paper, we propose a simple yet effective video super-resolution method that aims at generating highfidelity high-resolution (HR) videos from low-resolution (LR) ones. Previous methods predominantly leverage temporal neighbor frames to assist the super-resolution of the current frame. Those methods achieve limited performance as they suffer from the challenges in spatial frame alignment and the lack of useful information from similar LR neighbor frames. In contrast, we devise a cross-frame non-local attention mechanism that allows video superresolution without frame alignment, leading to being more robust to large motions in the video. In addition, to acquire general video prior information beyond neighbor frames, and to compensate for the information loss caused by large motions, we design a novel memory-augmented attention module to memorize general video details during the superresolution training. We have thoroughly evaluated our work on various challenging datasets. Compared to other recent video super-resolution approaches, our method not only achieves significant performance gains on large motion videos but also shows better generalization. Our source code and the new Parkour benchmark dataset is available at https://github.com/jiy173/MANA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c78159df-bcbc-4348-88da-67b22360bd0fCited by top-tier papers10
- Learning Non-Local Spatial-Angular Correlation for Light Field Image Super-ResolutionZhengyu Liang, Yingqian Wang, Longguang Wang, Jungang Yang et al.ICCV 2023 · 72 citations
- Sparse Instance Conditioned Multimodal Trajectory PredictionYonghao Dong, Le Wang, Sanping Zhou, Gang HuaICCV 2023 · 30 citations
- Learning Truncated Causal History Model for Video RestorationAmirhosein Ghasemabadi, Muhammad Kamran Janjua, Mohammad Salameh, Di NiuNeurIPS 2024 · 28 citations
- SAVSR: Arbitrary-Scale Video Super-Resolution via a Learned Scale-Adaptive NetworkZekun Li, Hongying Liu, Fanhua Shang, Yuanyuan Liu et al.AAAI 2024 · 23 citations
- Driving-Video Dehazing with Non-Aligned Regularization for Safety AssistanceJunkai Fan, Jiangwei Weng, Kun Wang, Yijun Yang et al.CVPR 2024
Builds on11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Progressive Fusion Video Super-Resolution Network via Exploiting Non-Local Spatio-Temporal CorrelationsPeng Yi, Zhongyuan Wang, Kui Jiang, Junjun Jiang et al.ICCV 2019 · 309 citations
- Deep Blind Video Super-resolutionJinshan Pan, Haoran Bai, Jiangxin Dong, Jiawei Zhang et al.ICCV 2021 · 65 citations
- Progressive Memory Banks for Incremental Domain AdaptationNabiha Asghar, Lili Mou, Kira A. Selby, Kevin D. Pantasdo et al.ICLR 2020 · 26 citations
- TDAN: Temporally-Deformable Alignment Network for Video Super-ResolutionYapeng Tian, Yulun Zhang, Yun Fu, Chenliang XuCVPR 2020
Related papers
- Video Super-Resolution With Temporal Group AttentionTakashi Isobe, Songjiang Li, Xu Jia, Shanxin Yuan et al.CVPR 2020
- Learning Trajectory-Aware Transformer for Video Super-ResolutionChengxu Liu, Huan Yang, Jianlong Fu, Xueming QianCVPR 2022 · 113 citations
- Stereo Video Super-Resolution via Exploiting View-Temporal CorrelationsRuikang Xu, Zeyu Xiao, Mingde Yao, Yueyi Zhang et al.ACM MM 2021 · 20 citations
- Image Super-Resolution With Cross-Scale Non-Local Attention and Exhaustive Self-Exemplars MiningYiqun Mei, Yuchen Fan, Yuqian Zhou, Lichao Huang et al.CVPR 2020
- Event-Enhanced Blurry Video Super-ResolutionDachun Kai, Yueyi Zhang, Jin Wang, Zeyu Xiao et al.AAAI 2025 · 7 citations
