Learning-Based Video Coding with Joint Deep Compression and Enhancement
Tiesong Zhao, Weize Feng, Hongji Zeng, Yiwen Xu, Yuzhen Niu, Jiaying Liu
Abstract
The end-to-end learning-based video compression has attracted substantial attentions by paving another way to compress video signals as stacked visual features. This paper proposes an efficient end-to-end deep video codec with jointly optimized compression and enhancement modules (JCEVC). First, we propose a dual-path generative adversarial network (DPEG) to reconstruct video details after compression. An 𝛼-path facilitates the structure information reconstruction with a large receptive field and multi-frame references, while a 𝛽-path facilitates the reconstruction of local textures. Both paths are fused and co-trained within a generative-adversarial process. Second, we reuse the DPEG network in both motion compensation and quality enhancement modules, which are further combined with other necessary modules to formulate our JCEVC framework. Third, we employ a joint training of deep video compression and enhancement that further improves the rate-distortion (RD) performance of compression. Compared with x265 LDP very fast mode, our JCEVC reduces the average bit-per-pixel (bpp) by 39.39%/54.92% at the same PSNR/MS-SSIM, which outperforms the state-of-the-art deep video codecs by a considerable margin.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- PDStream: Slashing Long- Tail Delay in Interactive Video Streaming via Pseudo-Dual StreamingXuedou Xiao, Yingying Zuo, Mingxuan Yan, Kezhong Liu et al.INFOCOM 2025 · 1 citation
- Perceptual Neural Video Compression with Color Separation and Rank Chainxiongzhuang liang, Chuanbo Tang, Zhuoyuan Li, Li Li et al.CVPR 2026
Builds on13
- Learned Video CompressionOren Rippel, Sanjay Nair, Carissa Lew, Steve Branson et al.ICCV 2019 · 258 citations
- Video Compression With Rate-Distortion AutoencodersAmirHossein Habibian, Ties van Rozendaal, Jakub M. Tomczak, Taco CohenICCV 2019 · 233 citations
- ELF-VC: Efficient Learned Flexible-Rate Video CodingOren Rippel, Alexander G. Anderson, Kedar Tatwawadi, Sanjay Nair et al.ICCV 2021 · 137 citations
- Online-trained Upsampler for Deep Low Complexity Video CompressionJan P. Klopp, Keng-Chi Liu, Shao-Yi Chien, Liang-Gee ChenICCV 2021 · 8 citations
- M-LVC: Multiple Frames Prediction for Learned Video CompressionJianping Lin, Dong Liu, Houqiang Li, Feng WuCVPR 2020
Related papers
- Learning for Video Compression With Hierarchical Quality and Recurrent EnhancementRen Yang, Fabian Mentzer, Luc Van Gool, Radu TimofteCVPR 2020
- Neural Inter-Frame Compression for Video CodingAbdelaziz Djelouah, Joaquim Campos, Simone Schaub-Meyer, Christopher SchroersICCV 2019 · 207 citations
- Offline and Online Optical Flow Enhancement for Deep Video CompressionChuanbo Tang, Xihua Sheng, Zhuoyuan Li, Haotian Zhang et al.AAAI 2024 · 35 citations
- DCDiff: Enhancing JPEG Compression via Diffusion-based DC Coefficients EstimationZiyuan Zhang, Han Qiu, Tianwei Zhang, Bin Chen et al.DAC 2025
- Complexity-guided Slimmable Decoder for Efficient Deep Video CompressionZhihao Hu, Dong XuCVPR 2023
