Hierarchical Frequency-Guided Alignment Transformer for Compressed Video Quality Enhancement
Liuhan Peng, Shuai Li, Yanbo Gao, Mao Ye, Chong Lv
摘要
During the video encoding process, the original spatial domain signal is first transformed into the frequency domain, followed by quantization and compression. As a result, the quality degradation in compressed videos primarily stems from distortions in the frequency domain information. However, existing video enhancement methods typically directly fuse information from adjacent frames in the spatial domain, making it difficult for models to effectively compensate for frequency domain distortions, which leads to suboptimal detail restoration. To address this issue, we propose a Hierarchical Frequency-Guided Alignment Transformer. Additionally, by analyzing the characteristics of the frequency domain, we find that different frequency bands exhibit both correlations and a certain degree of independence. Based on this, we introduce a Frequency-Aware Transformer module that employs a combination of independent and mixed processing to optimize information exchange across different frequency domains, effectively mitigating cross-interference from irrelevant information. Experimental results demonstrate that, compared to existing methods, our approach achieves state-of-the-art performance in objective metrics (PSNR/SSIM), perceptual quality (LPIPS), and subjective visual effects, while reducing model complexity.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou 等CVPR 2022 · 被引用 1,970 次
- Spatio-Temporal Deformable Convolution for Compressed Video Quality EnhancementJianing Deng, Li Wang, Shiliang Pu, Cheng ZhuoAAAI 2020 · 被引用 168 次
- Recursive Fusion and Deformable Spatiotemporal Attention for Video Compression Artifact ReductionMinyi Zhao, Yi Xu, Shuigeng ZhouACM MM 2021 · 被引用 61 次
- Video Compression Artifact Reduction by Fusing Motion Compensation and Global Context in a Swin-CNN Based Parallel ArchitectureXinjian Zhang, Su Yang, Wuyang Luo, Longwen Gao 等AAAI 2023 · 被引用 15 次
- CPGA: Coding Priors-Guided Aggregation Network for Compressed Video Quality EnhancementQiang Zhu, Jinhua Hao, Yukang Ding, Yu Liu 等CVPR 2024 · 被引用 13 次
相关 Paper
- Frequency-Aware Spatiotemporal Transformers for Video Inpainting DetectionBingyao Yu, Wanhua Li, Xiu Li, Jiwen Lu 等ICCV 2021 · 被引用 38 次
- Exploring Temporal Frequency Spectrum in Deep Video DeblurringQi Zhu, Man Zhou, Naishan Zheng, Chongyi Li 等ICCV 2023 · 被引用 21 次
- Rethinking Diffusion Model-Based Video Super-Resolution: Leveraging Dense Guidance from Aligned FeaturesJingyi Xu, Meisong Zheng, Ying Chen, Minglang Qiao 等CVPR 2026 · 被引用 1 次
- DreamUHD: Frequency Enhanced Variational Autoencoder for Ultra-High-Definition Image RestorationYidi Liu, Dong Li, Jie Xiao, Yuanfei Bao 等AAAI 2025 · 被引用 11 次
- Video Frame Interpolation TransformerZhihao Shi, Xiangyu Xu, Xiaohong Liu, Jun Chen 等CVPR 2022 · 被引用 117 次
