Multi-Frequency Representation Enhancement with Privilege Information for Video Super-Resolution
Fei Li, Linfeng Zhang, Zikun Liu, Juan Lei, Zhenbo Li
Abstract
CNN’s limited receptive field restricts its ability to capture long-range spatial-temporal dependencies, leading to unsatisfactory performance in video super-resolution (VSR). To tackle this challenge, this paper presents a novel multi-frequency representation enhancement module (MFE) that performs spatial-temporal information aggregation in the frequency domain. Specifically, MFE mainly includes a spatial-frequency representation enhancement branch which captures the long-range dependency in the spatial dimension, and an energy frequency representation enhancement branch to obtain the inter-channel feature relationship. Moreover, a novel model training method named privilege training is proposed to encode the privilege information from high-resolution videos to facilitate model training. With these two methods, we introduce a new VSR model named MFPI, which outperforms state-of-the-art methods by a large margin while maintaining good efficiency on various datasets, including REDS4, Vimeo, Vid4, and UDM10.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c1bfe619-e8bc-4991-9092-5de7ed35b8b1Cited by top-tier papers4
- MambaOVSR: Multiscale Fusion with Global Motion Modeling for Chinese Opera Video Super-ResolutionHua Chang, Xin Xu, Wei Liu, Wei Wang et al.AAAI 2026 · 1 citation
- Event-based Video Super-Resolution via State Space ModelsZeyu Xiao, Xinchao WangCVPR 2025
- IQ-VFI: Implicit Quadratic Motion Estimation for Video Frame InterpolationMengshun Hu, Kui Jiang, Zhihang Zhong, Zheng Wang et al.CVPR 2024
- VideoGigaGAN: Towards Detail-rich Video Super-ResolutionYiran Xu, Taesung Park, Richard Zhang, Yang Zhou et al.CVPR 2025
Builds on26
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- SimAM: A Simple, Parameter-Free Attention Module for Convolutional Neural NetworksLingxiao Yang, Ru-Yuan Zhang, Lida Li, Xiaohua XieICML 2021 · 1,593 citations
- SegNeXt: Rethinking Convolutional Attention Design for Semantic SegmentationMeng-Hao Guo, Cheng-Ze Lu, Qibin Hou, Zhengning Liu et al.NeurIPS 2022 · 1,385 citations
- Be Your Own Teacher: Improve the Performance of Convolutional Neural Networks via Self DistillationLinfeng Zhang, Jiebo Song, Anni Gao, Jingwei Chen et al.ICCV 2019 · 1,069 citations
- FcaNet: Frequency Channel Attention NetworksZequn Qin, Pengyi Zhang, Fei Wu, Xi LiICCV 2021 · 1,049 citations
Related papers
- How Video Super-Resolution and Frame Interpolation Mutually BenefitChengcheng Zhou, Zongqing Lu, Linge Li, Qiangyu Yan et al.ACM MM 2021 · 12 citations
- Time Without Time: Pseudo-Temporal Representation for Space-Time Super-ResolutionHee Min Choi, Hyoa Kang, Suji Kim, Dokwan Oh et al.CVPR 2026
- LDIP: Long Distance Information Propagation for Video Super-ResolutionMichael Bernasconi, Abdelaziz Djelouah, Yang Zhang, Markus Gross et al.ICCV 2025 · 1 citation
- Frequency-Aware Spatiotemporal Transformers for Video Inpainting DetectionBingyao Yu, Wanhua Li, Xiu Li, Jiwen Lu et al.ICCV 2021 · 38 citations
- Video Super-Resolution With Temporal Group AttentionTakashi Isobe, Songjiang Li, Xu Jia, Shanxin Yuan et al.CVPR 2020
