Selective compression learning of latent representations for variable-rate image compression
Jooyoung Lee, Seyoon Jeong, Munchurl Kim
摘要
Recently, many neural network-based image compression methods have shown promising results superior to the existing tool-based conventional codecs. However, most of them are often trained as separate models for different target bit rates, thus increasing the model complexity. Therefore, several studies have been conducted for learned compression that supports variable rates with single models, but they require additional network modules, layers, or inputs that often lead to complexity overhead, or do not provide sufficient coding efficiency. In this paper, we firstly propose a selective compression method that partially encodes the latent representations in a fully generalized manner for deep learning-based variable-rate image compression. The proposed method adaptively determines essential representation elements for compression of different target quality levels. For this, we first generate a 3D importance map as the nature of input content to represent the underlying importance of the representation elements. The 3D importance map is then adjusted for different target quality levels using importance adjustment curves. The adjusted 3D importance map is finally converted into a 3D binary mask to determine the essential representation elements for compression. The proposed method can be easily integrated with the existing compression models with a negligible amount of overhead increase. Our method can also enable continuously variable-rate compression via simple interpolation of the importance adjustment curves among different quality levels. The extensive experimental results show that the proposed method can achieve comparable compression efficiency as those of the separately trained reference compression models and can reduce decoding time owing to the selective compression. The sample codes are publicly available at https://github.com/JooyoungLeeETRI/SCR .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- All-in-One Image Coding for Joint Human-Machine Vision with Multi-Path AggregationXu Zhang, Peiyao Guo, Ming Lu, Zhan MaNeurIPS 2024 · 被引用 20 次
- COMPASS: High-Efficiency Deep Image Compression with Arbitrary-scale Spatial ScalabilityJongmin Park, Jooyoung Lee, Munchurl KimICCV 2023 · 被引用 6 次
- 3D Gaussian Splatting Data Compression with Mixture of PriorsLei Liu, Zhenghao Chen, Dong XuACM MM 2025 · 被引用 4 次
- GLIC: General Format Learned Image CompressionMingsheng Zhou, Mingming KongAAAI 2025 · 被引用 1 次
- Once-for-All: Controllable Generative Image Compression with Dynamic Granularity AdaptationAnqi Li, Feng Li, Yuxi Liu, Runmin Cong 等ICLR 2025
它引用的顶会 Paper4
- Variable Rate Deep Image Compression With a Conditional AutoencoderYoojin Choi, Mostafa El-Khamy, Jungwon LeeICCV 2019 · 被引用 265 次
- ELF-VC: Efficient Learned Flexible-Rate Video CodingOren Rippel, Alexander G. Anderson, Kedar Tatwawadi, Sanjay Nair 等ICCV 2021 · 被引用 137 次
- Variable-Rate Deep Image Compression through Spatially-Adaptive Feature TransformMyungseo Song, Jinyoung Choi, Bohyung HanICCV 2021 · 被引用 129 次
- Learned Image Compression With Discretized Gaussian Mixture Likelihoods and Attention ModulesZhengxue Cheng, Heming Sun, Masaru Takeuchi, Jiro KattoCVPR 2020
相关 Paper
- High-Fidelity Variable-Rate Image Compression via Invertible Activation TransformationShilv Cai, Zhijun Zhang, Liqun Chen, Luxin Yan 等ACM MM 2022 · 被引用 16 次
- Video Compression With Rate-Distortion AutoencodersAmirHossein Habibian, Ties van Rozendaal, Jakub M. Tomczak, Taco CohenICCV 2019 · 被引用 233 次
- EVC: Towards Real-Time Neural Image Compression with Mask DecayGuo-Hua Wang, Jiahao Li, Bin Li, Yan LuICLR 2023 · 被引用 24 次
- Interpolation Variable Rate Image CompressionZhenhong Sun, Zhiyu Tan, Xiuyu Sun, Fangyi Zhang 等ACM MM 2021 · 被引用 18 次
- Asymmetric Gained Deep Image Compression With Continuous Rate AdaptationZe Cui, Jing Wang, Shangyin Gao, Tiansheng Guo 等CVPR 2021
