Deep Contextual Video Compression
Jiahao Li, Bin Li, Yan Lu
Abstract
Most of the existing neural video compression methods adopt the predictive coding framework, which first generates the predicted frame and then encodes its residue with the current frame. However, as for compression ratio, predictive coding is only a sub-optimal solution as it uses simple subtraction operation to remove the redundancy across frames. In this paper, we propose a deep contextual video compression framework to enable a paradigm shift from predictive coding to conditional coding. In particular, we try to answer the following questions: how to define, use, and learn condition under a deep video compression framework. To tap the potential of conditional coding, we propose using feature domain context as condition. This enables us to leverage the high dimension context to carry rich information to both the encoder and the decoder, which helps reconstruct the high-frequency contents for higher video quality. Our framework is also extensible, in which the condition can be flexibly designed. Experiments show that our method can significantly outperform the previous state-of-the-art (SOTA) deep video compression methods. When compared with x265 using veryslow preset, we can achieve 26.0% bitrate saving for 1080P standard test videos.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ba03470d-ad31-4430-b2a5-fb049ca0959fCited by top-tier papers89
- Hybrid Spatial-Temporal Entropy Modelling for Neural Video CompressionJiahao Li, Bin Li, Yan LuACM MM 2022 · 202 citations
- VCT: A Video Compression TransformerFabian Mentzer, George Toderici, David Minnen, Sergi Caelles et al.NeurIPS 2022 · 155 citations
- HiNeRV: Video Compression with Hierarchical Encoding-based Neural RepresentationHo Man Kwan, Ge Gao, Fan Zhang, Andrew Gower et al.NeurIPS 2023 · 132 citations
- Coarse-To-Fine Deep Video Coding with Hyperprior-Guided Mode PredictionZhihao Hu, Guo Lu, Jinyang Guo, Shan Liu et al.CVPR 2022 · 95 citations
- CompGS: Efficient 3D Scene Representation via Compressed Gaussian SplattingXiangrui Liu, Xinju Wu, Pingping Zhang, Shiqi Wang et al.ACM MM 2024 · 52 citations
Builds on9
- Learned Video CompressionOren Rippel, Sanjay Nair, Carissa Lew, Steve Branson et al.ICCV 2019 · 258 citations
- Video Compression With Rate-Distortion AutoencodersAmirHossein Habibian, Ties van Rozendaal, Jakub M. Tomczak, Taco CohenICCV 2019 · 233 citations
- Neural Inter-Frame Compression for Video CodingAbdelaziz Djelouah, Joaquim Campos, Simone Schaub-Meyer, Christopher SchroersICCV 2019 · 207 citations
- Hierarchical Autoregressive Modeling for Neural Video CompressionRuihan Yang, Yibo Yang, Joseph Marino, Stephan MandtICLR 2021 · 48 citations
- M-LVC: Multiple Frames Prediction for Learned Video CompressionJianping Lin, Dong Liu, Houqiang Li, Feng WuCVPR 2020
Related papers
- Extending Neural P-frame Codecs for B-frame CodingReza Pourreza, Taco CohenICCV 2021 · 52 citations
- Neural Video Compression with In-Loop Contextual Filtering and Out-of-Loop Reconstruction EnhancementYaojun Wu, Chaoyi Lin, Yiming Wang, Semih Esenlik et al.ACM MM 2025 · 1 citation
- Neural Video Compression with Context ModulationChuanbo Tang, Zhuoyuan Li, Yifan Bian, Li Li et al.CVPR 2025
- Neural Video Compression with Feature ModulationJiahao Li, Bin Li, Yan LuCVPR 2024
- Neural Video Compression with Diverse ContextsJiahao Li, Bin Li, Yan LuCVPR 2023
