Discernible Image Compression
Zhaohui Yang, Yunhe Wang, Chang Xu, Peng Du, Chao Xu, Chunjing Xu, Qi Tian
Abstract
Image compression, as one of the fundamental low-level image processing tasks, is very essential for computer vision. Tremendous computing and storage resources can be preserved with a trivial amount of visual information. Conventional image compression methods tend to obtain compressed images by minimizing their appearance discrepancy with the corresponding original images, but pay little attention to their efficacy in downstream perception tasks, e.g., image recognition and object detection. Thus, some of compressed images could be recognized with bias. In contrast, this paper aims to produce compressed images by pursuing both appearance and perceptual consistency. Based on the encoder-decoder framework, we propose using a pre-trained CNN to extract features of the original and compressed images, and making them similar. Thus the compressed images are discernible to subsequent tasks, and we name our method as Discernible Image Compression (DIC). In addition, the maximum mean discrepancy (MMD) is employed to minimize the difference between feature distributions. The resulting compression network can generate images with high image quality and preserve the consistent perception in the feature domain, so that these images can be well recognized by pre-trained machine learning models. Experiments on benchmarks demonstrate that images compressed by using the proposed method can also be well recognized by subsequent visual recognition and detection models. For instance, the mAP value of compressed images by DIC is about 0.6% higher than that of using compressed images by conventional methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f832d78f-9568-4b9a-8a33-1c4101ed11e1Cited by top-tier papers4
- Non-Semantics Suppressed Mask Learning for Unsupervised Video Semantic CompressionYuan Tian, Guo Lu, Guangtao Zhai, Zhiyong GaoICCV 2023 · 29 citations
- Discovering Adaptive Task Dependencies for Efficient Multi-Task Representation CompressionZhimeng Huang, Rongao Yuan, Junlong Gao, Qi Mao et al.CVPR 2026
- Image Quality Assessment: From Human to Machine PreferenceChunyi Li, Yuan Tian, Xiaoyue Ling, Zicheng Zhang et al.CVPR 2025
- HourNAS: Extremely Fast Neural Architecture Search Through an Hourglass LensZhaohui Yang, Yunhe Wang, Xinghao Chen, Jianyuan Guo et al.CVPR 2021
Builds on6
- Data-Free Learning of Student NetworksHanting Chen, Yunhe Wang, Chang Xu, Zhaohui Yang et al.ICCV 2019 · 427 citations
- Variable Rate Deep Image Compression With a Conditional AutoencoderYoojin Choi, Mostafa El-Khamy, Jungwon LeeICCV 2019 · 265 citations
- Efficient Residual Dense Block Search for Image Super-ResolutionDehua Song, Chang Xu, Xu Jia, Yiyi Chen et al.AAAI 2020 · 145 citations
- CARS: Continuous Evolution for Efficient Neural Architecture SearchZhaohui Yang, Yunhe Wang, Xinghao Chen, Boxin Shi et al.CVPR 2020
- HourNAS: Extremely Fast Neural Architecture Search Through an Hourglass LensZhaohui Yang, Yunhe Wang, Xinghao Chen, Jianyuan Guo et al.CVPR 2021
Related papers
- Diff-ICMH: Harmonizing Machine and Human Vision in Image Compression with Generative PriorRuoyu Feng, Yunpeng Qi, Jinming Liu, Yixin Gao et al.NeurIPS 2025 · 5 citations
- SDiD:Shared diffusion prior for efficient distributed stereo image compressionYichong Xia, Yimin Zhou, Zongyu Li, Shiyu Qin et al.ICML 2026
- The Last Byte: Learning Just Enough for Machine-Oriented Image CompressionWuyuan Xie, Zhenming Li, Ye Liu, Jian Jin et al.AAAI 2026
- Understanding and Simplifying Perceptual DistancesDan Amir, Yair WeissCVPR 2021
- DiffPC: Diffusion-based High Perceptual Fidelity Image Compression with Semantic RefinementYichong Xia, Yimin Zhou, Jinpeng Wang, Baoyi An et al.ICLR 2025
