Video Compression with Entropy-Constrained Neural Representations
Carlos Gomes, Roberto Azevedo, Christopher Schroers
Abstract
Encoding videos as neural networks is a recently proposed approach that allows new forms of video processing. However, traditional techniques still outperform such neural video representation (NVR) methods for the task of video compression. This performance gap can be explained by the fact that current NVR methods: i) use architectures that do not efficiently obtain a compact representation of temporal and spatial information; and ii) minimize rate and distortion disjointly (first overfitting a network on a video and then using heuristic techniques such as post-training quantization or weight pruning to compress the model). We propose a novel convolutional architecture for video representation that better represents spatio-temporal information and a training strategy capable of jointly optimizing rate and distortion. All network and quantization parameters are jointly learned end-to-end, and the post-training operations used in previous works are unnecessary. We evaluate our method on the UVG dataset, achieving new state-ofthe-art results for video compression with NVRs. Moreover, we deliver the first NVR-based video compression method that improves over the typically adopted HEVC benchmark (x265, disabled b-frames, "medium" preset), closing the gap to autoencoder-based video compression techniques.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 514d8355-bb8f-42f3-929d-5973746f903aCited by top-tier papers8
- NVRC: Neural Video Representation CompressionHo Man Kwan, Ge Gao, Fan Zhang, Andrew Gower et al.NeurIPS 2024 · 44 citations
- PNVC: Towards Practical INR-based Video CompressionGe Gao, Ho Man Kwan, Fan Zhang, David BullAAAI 2025 · 20 citations
- Boosting Neural Representations for Videos with a Conditional DecoderXinjie Zhang, Ren Yang, Dailan He, Xingtong Ge et al.CVPR 2024 · 20 citations
- GIViC: Generative Implicit Video CompressionGe Gao, Siyue Teng, Tianhao Peng, Fan Zhang et al.ICCV 2025 · 4 citations
- Context Guided Transformer Entropy Modeling for Video CompressionJunlong Tong, Wei Zhang, Yaohui Jin, Xiaoyu ShenICCV 2025
Builds on6
- Learned Step Size quantizationSteven K. Esser, Jeffrey L. McKinstry, Deepika Bablani, Rathinakumar Appuswamy et al.ICLR 2020 · 1,037 citations
- High-Fidelity Generative Image CompressionFabian Mentzer, George Toderici, Michael Tschannen, Eirikur AgustssonNeurIPS 2020 · 675 citations
- Neural Inter-Frame Compression for Video CodingAbdelaziz Djelouah, Joaquim Campos, Simone Schaub-Meyer, Christopher SchroersICCV 2019 · 207 citations
- VCT: A Video Compression TransformerFabian Mentzer, George Toderici, David Minnen, Sergi Caelles et al.NeurIPS 2022 · 155 citations
- Scalable Model Compression by Entropy Penalized ReparameterizationDeniz Oktay, Johannes Ballé, Saurabh Singh, Abhinav ShrivastavaICLR 2020 · 46 citations
Related papers
- NIRVANA: Neural Implicit Representations of Videos with Adaptive Networks and Autoregressive Patch-Wise ModelingShishira R. Maiya, Sharath Girish, Max Ehrlich, Hanyu Wang et al.CVPR 2023
- Combining Frame and GOP Embeddings for Neural Video RepresentationJens Eirik Saethre, Roberto Azevedo, Christopher SchroersCVPR 2024
- NeRV: Neural Representations for VideosHao Chen, Bo He, Hanyu Wang, Yixuan Ren et al.NeurIPS 2021 · 430 citations
- FFNeRV: Flow-Guided Frame-Wise Neural Representations for VideosJoo Chan Lee, Daniel Rho, Jong Hwan Ko, Eunbyung ParkACM MM 2023 · 59 citations
- VRVVC: Variable-Rate NeRF-Based Volumetric Video CompressionQiang Hu, Houqiang Zhong, Zihan Zheng, Xiaoyun Zhang et al.AAAI 2025 · 11 citations
