Consistency Guided Diffusion Model with Neural Syntax for Perceptual Image Compression
Haowei Kuang, Yiyang Ma, Wenhan Yang, Zongming Guo, Jiaying Liu
Abstract
Diffusion models show impressive performances in image generation with excellent perceptual quality. However, its tendency to introduce additional distortion prevents its direct application in image compression. To address the issue, this paper introduces a Consistency Guided Diffusion Model (CGDM) tailored for perceptual image compression, which integrates an end-to-end image compression model with a diffusion-based post-processing network, aiming to learn richer detail representations with less fidelity loss. In detail, the compression and post-processing networks are cascaded and a branch of consistency guided features is added to constrain the deviation in the diffusion process for better reconstruction quality. Furthermore, a Syntax driven Feature Fusion (SFF) module is constructed to take an extra ultra-low bitstream from the encoding end as input, guiding the adaptive fusion of information from the two branches. In addition, we design a globally uniform boundary control strategy with overlapped patches and adopt a continuous online optimization mode to improve both coding efficiency and global consistency. Extensive experiments validate the superiority of our method to existing perceptual compression techniques. Our project is publicly available at: https://ellisonkuang.github.io/CGDM.github.io/.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers6
- Diff-ICMH: Harmonizing Machine and Human Vision in Image Compression with Generative PriorRuoyu Feng, Yunpeng Qi, Jinming Liu, Yixin Gao et al.NeurIPS 2025 · 5 citations
- CADC: Content Adaptive Diffusion-Based Generative Image CompressionXihua Sheng, Lingyu Zhu, Tianyu Zhang, Dong Liu et al.CVPR 2026 · 2 citations
- Cross-Granularity Online Optimization with Masked Compensated Information for Learned Image CompressionHaowei Kuang, Wenhan Yang, Zongming Guo, Jiaying LiuICCV 2025 · 1 citation
- Decouple Distortion from Perception: Region Adaptive Diffusion for Extreme-low Bitrate Perception Image CompressionJinchang Xu, Shaokang Wang, Jintao Chen, Zhe Li et al.CVPR 2025
- SSCL: Adversarially Guided Image Compression via Semantic and Spectral Consistency LearningWei Jiang, Yongqi Zhai, Jiayu Yang, Bohao Feng et al.AAAI 2026
Related papers
- Correcting Diffusion-Based Perceptual Image Compression with Privileged End-to-End DecoderYiyang Ma, Wenhan Yang, Jiaying LiuICML 2024 · 12 citations
- Lossy Image Compression with Conditional Diffusion ModelsRuihan Yang, Stephan MandtNeurIPS 2023 · 268 citations
- DiffPC: Diffusion-based High Perceptual Fidelity Image Compression with Semantic RefinementYichong Xia, Yimin Zhou, Jinpeng Wang, Baoyi An et al.ICLR 2025
- Towards Efficient Low-rate Image Compression with Frequency-aware Diffusion Prior RefinementYichong Xia, Yimin Zhou, Jinpeng Wang, Bin ChenAAAI 2026 · 1 citation
- Ultra Lowrate Image Compression with Semantic Residual Coding and Compression-aware DiffusionAnle Ke, Xu Zhang, Tong Chen, Ming Lu et al.ICML 2025
