TRACE: A Fast Transformer-based General-Purpose Lossless Compressor
Yu Mao, Yufei Cui, Tei-Wei Kuo, Chun Jason Xue
Abstract
Deep-learning-based compressor has received interests recently due to much improved compression ratio. However, modern approaches suffer from long execution time. To ease this problem, this paper targets on cutting down the execution time of deep-learning-based compressors. Building historydependencies sequentially (e.g., recurrent neural networks) is responsible for long inference latency. Instead, we introduce transformer into deep learning compressors to build historydependencies in parallel. However, existing transformer is too heavy in computation and incompatible to compression tasks. This paper proposes a fast general-purpose lossless compressor, TRACE, by designing a compression-friendly structure based on a single-layer transformer. We first design a new metric to advise the selection part of compression model structures. Byte-grouping and Shared-ffn schemes are further proposed to fully utilize the capacity of the singlelayer transformer. These features allow TRACE to achieve competitive compression ratio and a much faster speed. In addition, we further accelerate the compression procedure by designing a controller to reduce the parameter updating overhead. Experiments show that TRACE achieves an overall ∼3x speedup while keeps a comparable compression ratio to the state-of-the-art compressors. The source code for TRACE and links to the datasets are available at https:// github.com/mynotwo/A-Fast-Transformer-based-General-Purpose- LosslessCompressor. CCS CONCEPTS • Information systems → Data compression.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c848ff48-e740-4170-96f2-a437355d69d0Cited by top-tier papers11
- Language Modeling Is CompressionGrégoire Delétang, Anian Ruoss, Paul-Ambroise Duquenne, Elliot Catt et al.ICLR 2024 · 243 citations
- Faster and Stronger Lossless Compression with Optimized Autoregressive FrameworkYu Mao, Jingzong Li, Yufei Cui, Chun Jason XueDAC 2023 · 12 citations
- Ariadne: A Hotness-Aware and Size-Adaptive Compressed Swap Technique for Fast Application Relaunch and Reduced CPU Usage on Mobile DevicesYu Liang, Aofeng Shen, Chun Jason Xue, Riwei Pan et al.HPCA 2025 · 8 citations
- Genomics Data Lossless Compression with (S, K)-Mer Encoding and Deep Neural NetworksHui Sun, Liping Yi, Huidong Ma, Yongxia Sun et al.AAAI 2025 · 2 citations
- EDPC: Accelerating Lossless Compression via Lightweight Probability Models and Decoupled Parallel DataflowZeyi Lu, Xiaoxiao Ma, Yujun Huang, Minxiao Chen et al.ACM MM 2025 · 1 citation
Builds on6
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- Long Range Arena : A Benchmark for Efficient TransformersYi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen et al.ICLR 2021 · 881 citations
- Compressive Transformers for Long-Range Sequence ModellingJack W. Rae, Anna Potapenko, Siddhant M. Jayakumar, Chloe Hillier et al.ICLR 2020 · 833 citations
- IDF++: Analyzing and Improving Integer Discrete Flows for Lossless CompressionRianne van den Berg, Alexey A. Gritsenko, Mostafa Dehghani, Casper Kaae Sønderby et al.ICLR 2021 · 38 citations
- Two-Level Data Compression using Machine Learning in Time Series DatabaseXinyang Yu, Yanqing Peng, Feifei Li, Sheng Wang et al.ICDE 2020 · 36 citations
Related papers
- Accelerating General-purpose Lossless Compression via Simple and Scalable ParameterizationYu Mao, Yufei Cui, Tei-Wei Kuo, Chun Jason XueACM MM 2022 · 8 citations
- Temporal Latent Bottleneck: Synthesis of Fast and Slow Processing Mechanisms in Sequence LearningAniket Didolkar, Kshitij Gupta, Anirudh Goyal, Nitesh B. Gundavarapu et al.NeurIPS 2022 · 24 citations
- MSDZip: Universal Lossless Compression for Multi-source Data via Stepwise-parallel and Learning-based PredictionHuidong Ma, Hui Sun, Liping Yi, Yanfeng Ding et al.WWW 2025 · 9 citations
- Reducing the GPU Memory Bottleneck with Lossless Compression for MLAditya K. Kamath, Arvind Krishnamurthy, Marco Canini, Simon PeterEuroSys 2026
- Fast Lossless Neural Compression with Integer-Only Discrete FlowsSiyu Wang, Jianfei Chen, Chongxuan Li, Jun Zhu et al.ICML 2022 · 8 citations
