Variable-Rate Deep Image Compression through Spatially-Adaptive Feature Transform
Myungseo Song, Jinyoung Choi, Bohyung Han
Abstract
We propose a versatile deep image compression network based on Spatial Feature Transform (SFT) [47] , which takes a source image and a corresponding quality map as inputs and produce a compressed image with variable rates. Our model covers a wide range of compression rates using a single model, which is controlled by arbitrary pixel-wise quality maps. In addition, the proposed framework allows us to perform task-aware image compressions for various tasks, e.g., classification, by efficiently estimating optimized quality maps specific to target tasks for our encoding network. This is even possible with a pretrained network without learning separate models for individual tasks. Our algorithm achieves outstanding rate-distortion trade-off compared to the approaches based on multiple models that are optimized separately for several different target rates. At the same level of compression, the proposed approach successfully improves performance on image classification and text region quality preservation via task-aware quality map estimation without additional model training. The code is available at the project website 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext da003b2c-a3e5-4d5e-9bc9-574e72d88f27Cited by top-tier papers14
- TransTIC: Transferring Transformer-based Image Compression from Human Perception to Machine PerceptionYi-Hsin Chen, Ying-Chieh Weng, Chia-Hao Kao, Cheng Chien et al.ICCV 2023 · 59 citations
- Selective compression learning of latent representations for variable-rate image compressionJooyoung Lee, Seyoon Jeong, Munchurl KimNeurIPS 2022 · 43 citations
- Flexible Neural Image Compression via Code EditingChenjian Gao, Tongda Xu, Dailan He, Yan Wang et al.NeurIPS 2022 · 34 citations
- All-in-One Image Coding for Joint Human-Machine Vision with Multi-Path AggregationXu Zhang, Peiyao Guo, Ming Lu, Zhan MaNeurIPS 2024 · 20 citations
- Boosting Neural Representations for Videos with a Conditional DecoderXinjie Zhang, Ren Yang, Dailan He, Xingtong Ge et al.CVPR 2024 · 20 citations
Builds on4
- High-Fidelity Generative Image CompressionFabian Mentzer, George Toderici, Michael Tschannen, Eirikur AgustssonNeurIPS 2020 · 675 citations
- Generative Adversarial Networks for Extreme Learned Image CompressionEirikur Agustsson, Michael Tschannen, Fabian Mentzer, Radu Timofte et al.ICCV 2019 · 648 citations
- Variable Rate Deep Image Compression With a Conditional AutoencoderYoojin Choi, Mostafa El-Khamy, Jungwon LeeICCV 2019 · 265 citations
- Transfer Learning From Synthetic to Real-Noise Denoising With Adaptive Instance NormalizationYoonsik Kim, Jae Woong Soh, Gu Yong Park, Nam Ik ChoCVPR 2020
Related papers
- JPEG Inspired Deep LearningAhmed H. Salamah, Kaixiang Zheng, Yiwen Liu, En-Hui YangICLR 2025
- High-Fidelity Variable-Rate Image Compression via Invertible Activation TransformationShilv Cai, Zhijun Zhang, Liqun Chen, Luxin Yan et al.ACM MM 2022 · 16 citations
- Enhanced Invertible Encoding for Learned Image CompressionYueqi Xie, Ka Leong Cheng, Qifeng ChenACM MM 2021 · 195 citations
- Entroformer: A Transformer-based Entropy Model for Learned Image CompressionYichen Qian, Xiuyu Sun, Ming Lin, Zhiyu Tan et al.ICLR 2022 · 194 citations
- Practical Learned Lossless JPEG Recompression with Multi-Level Cross-Channel Entropy Model in the DCT DomainLina Guo, Xinjie Shi, Dailan He, Yuanyuan Wang et al.CVPR 2022 · 8 citations
