MLIC: Multi-Reference Entropy Model for Learned Image Compression
Wei Jiang, Jiayu Yang, Yongqi Zhai, Peirong Ning, Feng Gao, Ronggang Wang
Abstract
Recently, learned image compression has achieved remarkable performance. The entropy model, which estimates the distribution of the latent representation, plays a crucial role in boosting rate-distortion performance. However, most entropy models only capture correlations in one dimension, while the latent representation contains channel-wise, local spatial, and global spatial correlations. To tackle this issue, we propose the Multi-Reference Entropy Model (MEM) and the advanced version, MEM+. These models capture the different types of correlations present in latent representation. Specifically, we first divide the latent representation into slices. When decoding the current slice, we use previously decoded slices as context and employ the attention map of the previously decoded slice to predict global correlations in the current slice. To capture local contexts, we introduce two enhanced checkerboard context capturing techniques that avoids performance degradation. Based on MEM and MEM+, we propose image compression models MLIC and MLIC+. Extensive experimental evaluations demonstrate that our MLIC and MLIC+ models achieve state-of-the-art performance, reducing BD-rate by 8.05% and 11.39% on the Kodak dataset compared to VTM-17.0 when measured in PSNR.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c4fc522c-5954-4f1b-8caf-7d65dff8be17Cited by top-tier papers44
- CompGS: Efficient 3D Scene Representation via Compressed Gaussian SplattingXiangrui Liu, Xinju Wu, Pingping Zhang, Shiqi Wang et al.ACM MM 2024 · 52 citations
- One-Step Diffusion-Based Image Compression with Semantic DistillationNaifu Xue, Zhaoyang Jia, Jiahao Li, Bin Li et al.NeurIPS 2025 · 28 citations
- End-to-End RGB-D Image Compression via Exploiting Channel-Modality RedundancyHuiming Zheng, Wei GaoAAAI 2024 · 15 citations
- Learning Optimal Lattice Vector Quantizers for End-to-end Neural Image CompressionXi Zhang, Xiaolin WuNeurIPS 2024 · 12 citations
- An End-to-End Real-World Camera Imaging PipelineKepeng Xu, Zijia Ma, Li Xu, Gang He et al.ACM MM 2024 · 9 citations
Builds on18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- ELIC: Efficient Learned Image Compression with Unevenly Grouped Space-Channel Contextual Adaptive CodingDailan He, Ziming Yang, Weikun Peng, Rui Ma et al.CVPR 2022 · 363 citations
- The Devil Is in the Details: Window-based Attention for Image CompressionRenjie Zou, Chunfeng Song, Zhaoxiang ZhangCVPR 2022 · 260 citations
- Transformer-based Transform CodingYinhao Zhu, Yang Yang, Taco CohenICLR 2022 · 218 citations
- Enhanced Invertible Encoding for Learned Image CompressionYueqi Xie, Ka Leong Cheng, Qifeng ChenACM MM 2021 · 195 citations
Related papers
- Learning Accurate Entropy Model with Global Reference for Image CompressionYichen Qian, Zhiyu Tan, Xiuyu Sun, Ming Lin et al.ICLR 2021 · 93 citations
- Learned Image Compression with Dictionary-based Entropy ModelJingbo Lu, Leheng Zhang, Xingyu Zhou, Mu Li et al.CVPR 2025
- Towards Efficient Image Compression Without Autoregressive ModelsMuhammad Salman Ali, Yeongwoong Kim, Maryam Qamar, Sung-Chang Lim et al.NeurIPS 2023 · 17 citations
- Efficient Learned Image Compression without Entropy CodingHao Cao, Wenqi Guo, Zhijin Qin, Jungong HanICML 2026
- Joint Global and Local Hierarchical Priors for Learned Image CompressionJun-Hyuk Kim, Byeongho Heo, Jong-Seok LeeCVPR 2022 · 98 citations
