High Visual-Fidelity Learned Video Compression
Meng Li, Yibo Shi, Jing Wang, Yunqi Huang
Abstract
With the growing demand for video applications, many advanced learned video compression methods have been developed, outperforming traditional methods in terms of objective quality metrics such as PSNR. Existing methods primarily focus on objective quality but tend to overlook perceptual quality. Directly incorporating perceptual loss into a learned video compression framework is non-trivial and raises several perceptual quality issues that need to be addressed. In this paper, we investigated these issues in learned video compression and propose a novel High Visual-Fidelity Learned Video Compression framework (HVFVC). Specifically, we design a novel confidence-based feature reconstruction method to address the issue of poor reconstruction in newly-emerged regions, which significantly improves the visual quality of the reconstruction. Furthermore, we present a periodic compensation loss to mitigate the checkerboard artifacts related to deconvolution operation and optimization. Extensive experiments have shown that the proposed HVFVC achieves excellent perceptual quality, outperforming the latest VVC standard with only 50% required bitrate.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 253546de-8ec6-4cdf-a5a5-2740e87a1b21Cited by top-tier papers4
- Single-step Diffusion-based Video Coding with Semantic-Temporal GuidanceNaifu Xue, Zhaoyang Jia, Jiahao Li, Bin Li et al.CVPR 2026 · 12 citations
- Perceptual Neural Video Compression with Color Separation and Rank Chainxiongzhuang liang, Chuanbo Tang, Zhuoyuan Li, Li Li et al.CVPR 2026
- Neural Video Compression with Reference HierarchyChuanbo Tang, Zhuoyuan Li, Li Li, Dong Liu et al.AAAI 2026
- Neural Video Compression with Context ModulationChuanbo Tang, Zhuoyuan Li, Yifan Bian, Li Li et al.CVPR 2025
Builds on9
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 2,878 citations
- High-Fidelity Generative Image CompressionFabian Mentzer, George Toderici, Michael Tschannen, Eirikur AgustssonNeurIPS 2020 · 675 citations
- Deep Contextual Video CompressionJiahao Li, Bin Li, Yan LuNeurIPS 2021 · 518 citations
- KNN Local Attention for Image RestorationHunsang Lee, Hyesong Choi, Kwanghoon Sohn, Dongbo MinCVPR 2022 · 62 citations
- Extending Neural P-frame Codecs for B-frame CodingReza Pourreza, Taco CohenICCV 2021 · 52 citations
Related papers
- FVC: A New Framework Towards Deep Video Compression in Feature SpaceZhihao Hu, Guo Lu, Dong XuCVPR 2021
- ELF-VC: Efficient Learned Flexible-Rate Video CodingOren Rippel, Alexander G. Anderson, Kedar Tatwawadi, Sanjay Nair et al.ICCV 2021 · 137 citations
- Learned Bi-Resolution Image Coding using Generalized Octave ConvolutionsMohammad Akbari, Jie Liang, Jingning Han, Chengjie TuAAAI 2021 · 21 citations
- Learned Video CompressionOren Rippel, Sanjay Nair, Carissa Lew, Steve Branson et al.ICCV 2019 · 258 citations
- Another Way to the Top: Exploit Contextual Clustering in Learned Image CodingYichi Zhang, Zhihao Duan, Ming Lu, Dandan Ding et al.AAAI 2024 · 13 citations
