Entroformer: A Transformer-based Entropy Model for Learned Image Compression
Yichen Qian, Xiuyu Sun, Ming Lin, Zhiyu Tan, Rong Jin
摘要
One critical component in lossy deep image compression is the entropy model, which predicts the probability distribution of the quantized latent representation in the encoding and decoding modules. Previous works build entropy models upon convolutional neural networks which are inefficient in capturing global dependencies. In this work, we propose a novel transformer-based entropy model, termed Entroformer, to capture long-range dependencies in probability distribution estimation effectively and efficiently. Different from vision transformers in image classification, the Entroformer is highly optimized for image compression, including a top-k self-attention and a diamond relative position encoding. Meanwhile, we further expand this architecture with a parallel bidirectional context model to speed up the decoding process. The experiments show that the Entroformer achieves state-of-the-art performance on image compression while being time-efficient. Code is available at https://github.com/damo-cv/entroformer .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper43
- CDTrans: Cross-domain Transformer for Unsupervised Domain AdaptationTongkun Xu, Weihua Chen, Pichao Wang, Fan Wang 等ICLR 2022 · 被引用 293 次
- VCT: A Video Compression TransformerFabian Mentzer, George Toderici, David Minnen, Sergi Caelles 等NeurIPS 2022 · 被引用 155 次
- MLIC: Multi-Reference Entropy Model for Learned Image CompressionWei Jiang, Jiayu Yang, Yongqi Zhai, Peirong Ning 等ACM MM 2023 · 被引用 117 次
- Frequency-Aware Transformer for Learned Image CompressionHan Li, Shaohui Li, Wenrui Dai, Chenglin Li 等ICLR 2024 · 被引用 88 次
- Causal Context Adjustment Loss for Learned Image CompressionMinghao Han, Shiyin Jiang, Shengxi Li, Xin Deng 等NeurIPS 2024 · 被引用 31 次
它引用的顶会 Paper5
- Coarse-to-Fine Hyper-Prior Modeling for Learned Image CompressionYueyu Hu, Wenhan Yang, Jiaying LiuAAAI 2020 · 被引用 143 次
- Learning Accurate Entropy Model with Global Reference for Image CompressionYichen Qian, Zhiyu Tan, Xiuyu Sun, Ming Lin 等ICLR 2021 · 被引用 93 次
- Ada-NETS: Face Clustering via Adaptive Neighbour Discovery in the Structure SpaceYaohua Wang, Yaobin Zhang, Fangyi Zhang, Senzhang Wang 等ICLR 2022 · 被引用 38 次
- Checkerboard Context Model for Efficient Learned Image CompressionDailan He, Yaoyan Zheng, Baocheng Sun, Yan Wang 等CVPR 2021
- Learned Image Compression With Discretized Gaussian Mixture Likelihoods and Attention ModulesZhengxue Cheng, Heming Sun, Masaru Takeuchi, Jiro KattoCVPR 2020
相关 Paper
- Joint Global and Local Hierarchical Priors for Learned Image CompressionJun-Hyuk Kim, Byeongho Heo, Jong-Seok LeeCVPR 2022 · 被引用 98 次
- Laplacian-guided Entropy Model in Neural Codec with Blur-dissipated SynthesisAtefeh Khoshkhahtinat, Ali Zafari, Piyush M. Mehta, Nasser M. NasrabadiCVPR 2024
- Learned Image Compression with Dictionary-based Entropy ModelJingbo Lu, Leheng Zhang, Xingyu Zhou, Mu Li 等CVPR 2025
- MIMT: Masked Image Modeling Transformer for Video CompressionJinxi Xiang, Kuan Tian, Jun ZhangICLR 2023
- Learned Image Compression with Mixed Transformer-CNN ArchitecturesJinming Liu, Heming Sun, Jiro KattoCVPR 2023
