TRACE: A Fast Transformer-based General-Purpose Lossless Compressor
Yu Mao, Yufei Cui, Tei-Wei Kuo, Chun Jason Xue
摘要
Deep-learning-based compressor has received interests recently due to much improved compression ratio. However, modern approaches suffer from long execution time. To ease this problem, this paper targets on cutting down the execution time of deep-learning-based compressors. Building historydependencies sequentially (e.g., recurrent neural networks) is responsible for long inference latency. Instead, we introduce transformer into deep learning compressors to build historydependencies in parallel. However, existing transformer is too heavy in computation and incompatible to compression tasks. This paper proposes a fast general-purpose lossless compressor, TRACE, by designing a compression-friendly structure based on a single-layer transformer. We first design a new metric to advise the selection part of compression model structures. Byte-grouping and Shared-ffn schemes are further proposed to fully utilize the capacity of the singlelayer transformer. These features allow TRACE to achieve competitive compression ratio and a much faster speed. In addition, we further accelerate the compression procedure by designing a controller to reduce the parameter updating overhead. Experiments show that TRACE achieves an overall ∼3x speedup while keeps a comparable compression ratio to the state-of-the-art compressors. The source code for TRACE and links to the datasets are available at https:// github.com/mynotwo/A-Fast-Transformer-based-General-Purpose- LosslessCompressor. CCS CONCEPTS • Information systems → Data compression.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Language Modeling Is CompressionGrégoire Delétang, Anian Ruoss, Paul-Ambroise Duquenne, Elliot Catt 等ICLR 2024 · 被引用 243 次
- Faster and Stronger Lossless Compression with Optimized Autoregressive FrameworkYu Mao, Jingzong Li, Yufei Cui, Chun Jason XueDAC 2023 · 被引用 12 次
- Ariadne: A Hotness-Aware and Size-Adaptive Compressed Swap Technique for Fast Application Relaunch and Reduced CPU Usage on Mobile DevicesYu Liang, Aofeng Shen, Chun Jason Xue, Riwei Pan 等HPCA 2025 · 被引用 8 次
- Genomics Data Lossless Compression with (S, K)-Mer Encoding and Deep Neural NetworksHui Sun, Liping Yi, Huidong Ma, Yongxia Sun 等AAAI 2025 · 被引用 2 次
- EDPC: Accelerating Lossless Compression via Lightweight Probability Models and Decoupled Parallel DataflowZeyi Lu, Xiaoxiao Ma, Yujun Huang, Minxiao Chen 等ACM MM 2025 · 被引用 1 次
它引用的顶会 Paper6
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- Long Range Arena : A Benchmark for Efficient TransformersYi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen 等ICLR 2021 · 被引用 881 次
- Compressive Transformers for Long-Range Sequence ModellingJack W. Rae, Anna Potapenko, Siddhant M. Jayakumar, Chloe Hillier 等ICLR 2020 · 被引用 833 次
- IDF++: Analyzing and Improving Integer Discrete Flows for Lossless CompressionRianne van den Berg, Alexey A. Gritsenko, Mostafa Dehghani, Casper Kaae Sønderby 等ICLR 2021 · 被引用 38 次
- Two-Level Data Compression using Machine Learning in Time Series DatabaseXinyang Yu, Yanqing Peng, Feifei Li, Sheng Wang 等ICDE 2020 · 被引用 36 次
相关 Paper
- Accelerating General-purpose Lossless Compression via Simple and Scalable ParameterizationYu Mao, Yufei Cui, Tei-Wei Kuo, Chun Jason XueACM MM 2022 · 被引用 8 次
- Temporal Latent Bottleneck: Synthesis of Fast and Slow Processing Mechanisms in Sequence LearningAniket Didolkar, Kshitij Gupta, Anirudh Goyal, Nitesh B. Gundavarapu 等NeurIPS 2022 · 被引用 24 次
- MSDZip: Universal Lossless Compression for Multi-source Data via Stepwise-parallel and Learning-based PredictionHuidong Ma, Hui Sun, Liping Yi, Yanfeng Ding 等WWW 2025 · 被引用 9 次
- Reducing the GPU Memory Bottleneck with Lossless Compression for MLAditya K. Kamath, Arvind Krishnamurthy, Marco Canini, Simon PeterEuroSys 2026
- Fast Lossless Neural Compression with Integer-Only Discrete FlowsSiyu Wang, Jianfei Chen, Chongxuan Li, Jun Zhu 等ICML 2022 · 被引用 8 次
