Semantic Scalable Image Compression with Cross-Layer Priors
Hanyue Tu, Li Li, Wengang Zhou, Houqiang Li
摘要
In an intelligent society, image compression needs to serve both human vision and machine vision. Traditional image compression schemes only consider visual quality for humans. In addition, the bitstream needs to be fully decoded to images before performing semantic analysis (e.g., by deep neural networks). These two factors make traditional image compression schemes semantically inefficient. To better serve the needs of both human vision and machine vision, it is more reasonable to compress and transmit image signals and features simultaneously. In this paper, we propose a novel end-to-end semantic scalable image compression method, which progressively compresses coarse-grained semantic features, fine-grained semantic features, and image signals. To utilize the cross-layer correlation between features and image signals, we propose a cross-layer context model to reduce the information redundancy, which takes higher-layer features as cross-layer priors to predict the probability distribution parameters for the entropy model of lower-layer features or images. Furthermore, we adopt a Region of Interest (ROI) compression scheme. The objects with rich semantic information and the background are compressed separately, to further improve the compression efficiency. Experimental results on the CUB-200-2011 and FGVC-Aircraft datasets demonstrate the effectiveness of our proposed scheme compared to separate compression of image signals and features.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- DeepSVC: Deep Scalable Video Coding for Both Machine and Human VisionHongbin Lin, Bolin Chen, Zhichen Zhang, Jielian Lin 等ACM MM 2023 · 被引用 32 次
- ICMH-Net: Neural Image Compression Towards both Machine Vision and Human VisionLei Liu, Zhihao Hu, Zhenghao Chen, Dong XuACM MM 2023 · 被引用 21 次
- Cross Modal Compression: Towards Human-comprehensible Semantic CompressionJiguo Li, Chuanmin Jia, Xinfeng Zhang, Siwei Ma 等ACM MM 2021 · 被引用 24 次
- Diff-ICMH: Harmonizing Machine and Human Vision in Image Compression with Generative PriorRuoyu Feng, Yunpeng Qi, Jinming Liu, Yixin Gao 等NeurIPS 2025 · 被引用 5 次
- ROI-Guided Point Cloud Geometry Compression Towards Human and Machine VisionLiang Xie, Wei Gao, Huiming Zheng, Ge LiACM MM 2024 · 被引用 51 次
