NVTC: Nonlinear Vector Transform Coding
Runsen Feng, Zongyu Guo, Weiping Li, Zhibo Chen
Abstract
In theory, vector quantization (VQ) is always better than scalar quantization (SQ) in terms of rate-distortion (R-D) performance [33] . Recent state-of-the-art methods for neural image compression are mainly based on nonlinear transform coding (NTC) with uniform scalar quantization, overlooking the benefits of VQ due to its exponentially increased complexity. In this paper, we first investigate on some toy sources, demonstrating that even if modern neural networks considerably enhance the compression performance of SQ with nonlinear transform, there is still an insurmountable chasm between SQ and VQ. Therefore, revolving around VQ, we propose a novel framework for neural image compression named Nonlinear Vector Transform Coding (NVTC). NVTC solves the critical complexity issue of VQ through (1) a multi-stage quantization strategy and (2) nonlinear vector transforms. In addition, we apply entropy-constrained VQ in latent space to adaptively determine the quantization boundaries for joint rate-distortion optimization, which improves the performance both theoretically and experimentally. Compared to previous NTC approaches, NVTC demonstrates superior rate-distortion performance, faster decoding speed, and smaller model size. Our code is available at https://github.com/ USTC-IMCL/NVTC.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f7defb4f-1166-41bb-aa3e-43d5973ac2a9Cited by top-tier papers6
- Learning Optimal Lattice Vector Quantizers for End-to-end Neural Image CompressionXi Zhang, Xiaolin WuNeurIPS 2024 · 12 citations
- Approaching Rate-Distortion Limits in Neural Compression with Lattice Transform CodingEric Lei, Hamed Hassani, Shirin Saeedi BidokhtiICLR 2025
- Balanced Rate-Distortion Optimization in Learned Image CompressionYichi Zhang, Zhihao Duan, Yuning Huang, Fengqing ZhuCVPR 2025
- Multirate Neural Image Compression with Adaptive Lattice Vector QuantizationHao Xu, Xiaolin Wu, Xi ZhangCVPR 2025
- Transform-Free Feature Coding via Entropy-Constrained Vector QuantizationQiaoxi Chen, Changsheng Gao, Li Li, Dong LiuAAAI 2026
Builds on8
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- The Devil Is in the Details: Window-based Attention for Image CompressionRenjie Zou, Chunfeng Song, Zhaoxiang ZhangCVPR 2022 · 260 citations
- Transformer-based Transform CodingYinhao Zhu, Yang Yang, Taco CohenICLR 2022 · 218 citations
- Improving Inference for Neural Image CompressionYibo Yang, Robert Bamler, Stephan MandtNeurIPS 2020 · 151 citations
Related papers
- SoftBinary Coding: A New Information-Theoretic Paradigm for Neural Compression via Fast Channel SimulationEzgi Ozyilkan, Sharang Sriramu, Elza Erkip, Aaron Wagner et al.ICML 2026
- Enhanced Invertible Encoding for Learned Image CompressionYueqi Xie, Ka Leong Cheng, Qifeng ChenACM MM 2021 · 195 citations
- LVQAC: Lattice Vector Quantization Coupled with Spatially Adaptive Companding for Efficient Learned Image CompressionXi Zhang, Xiaolin WuCVPR 2023
- EVC: Towards Real-Time Neural Image Compression with Mask DecayGuo-Hua Wang, Jiahao Li, Bin Li, Yan LuICLR 2023 · 24 citations
- Permute, Quantize, and Fine-Tune: Efficient Compression of Neural NetworksJulieta Martinez, Jashan Shewakramani, Ting-Wei Liu, Ioan Andrei Barsan et al.CVPR 2021
