Another Way to the Top: Exploit Contextual Clustering in Learned Image Coding
Yichi Zhang, Zhihao Duan, Ming Lu, Dandan Ding, Fengqing Zhu, Zhan Ma
摘要
While convolution and self-attention are extensively used in learned image compression (LIC) for transform coding, this paper proposes an alternative called Contextual Clustering based LIC (CLIC) which primarily relies on clustering operations and local attention for correlation characterization and compact representation of an image. As seen, CLIC expands the receptive field into the entire image for intra-cluster feature aggregation. Afterward, features are reordered to their original spatial positions to pass through the local attention units for inter-cluster embedding. Additionally, we introduce the Guided Post-Quantization Filtering (GuidedPQF) into CLIC, effectively mitigating the propagation and accumulation of quantization errors at the initial decoding stage. Extensive experiments demonstrate the superior performance of CLIC over state-of-the-art works: when optimized using MSE, it outperforms VVC by about 10% BD-Rate in three widely-used benchmark datasets; when optimized using MS-SSIM, it saves more than 50% BD-Rate over VVC. Our CLIC offers a new way to generate compact representations for image compression, which also provides a novel direction along the line of LIC development.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Content-Aware Mamba for Learned Image CompressionYunuo Chen, Zezheng Lyu, Bing He, Hongwei Hu 等ICLR 2026 · 被引用 5 次
- Adaptive Learned Image Compression with Graph Neural NetworksYunuo Chen, Bing He, Zezheng Lyu, Hongwei Hu 等CVPR 2026 · 被引用 1 次
- Balanced Rate-Distortion Optimization in Learned Image CompressionYichi Zhang, Zhihao Duan, Yuning Huang, Fengqing ZhuCVPR 2025
- Parameter-Free Clustering via Self-Supervised Consensus MaximizationLijun Zhang, Suyuan Liu, Siwei Wang, Shengju Yu 等AAAI 2026
它引用的顶会 Paper18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- MetaFormer is Actually What You Need for VisionWeihao Yu, Mi Luo, Pan Zhou, Chenyang Si 等CVPR 2022 · 被引用 1,114 次
- Differentiable Soft Quantization: Bridging Full-Precision and Low-Bit Neural NetworksRuihao Gong, Xianglong Liu, Shenghu Jiang, Tianxiang Li 等ICCV 2019 · 被引用 540 次
- ELIC: Efficient Learned Image Compression with Unevenly Grouped Space-Channel Contextual Adaptive CodingDailan He, Ziming Yang, Weikun Peng, Rui Ma 等CVPR 2022 · 被引用 363 次
- Variable Rate Deep Image Compression With a Conditional AutoencoderYoojin Choi, Mostafa El-Khamy, Jungwon LeeICCV 2019 · 被引用 265 次
相关 Paper
- Linear Attention Modeling for Learned Image CompressionDonghui Feng, Zhengxue Cheng, Shen Wang, Ronghua Wu 等CVPR 2025
- BiECVC: Gated Diversification of Bidirectional Contexts for Learned Video CompressionWei Jiang, Junru Li, Kai Zhang, Li ZhangACM MM 2025 · 被引用 3 次
- Learned Image Compression With Discretized Gaussian Mixture Likelihoods and Attention ModulesZhengxue Cheng, Heming Sun, Masaru Takeuchi, Jiro KattoCVPR 2020
- High Visual-Fidelity Learned Video CompressionMeng Li, Yibo Shi, Jing Wang, Yunqi HuangACM MM 2023 · 被引用 10 次
- Towards Efficient Image Compression Without Autoregressive ModelsMuhammad Salman Ali, Yeongwoong Kim, Maryam Qamar, Sung-Chang Lim 等NeurIPS 2023 · 被引用 17 次
