OmniZip: Learning a Unified and Lightweight Lossless Compressor for Multi-Modal Data
Yan Zhao, Zhengxue Cheng, Junxuan Zhang, Dajiang Zhou, Qunshan Gu, Qi Wang, Li Song
摘要
Lossless compression is essential for efficient data storage and transmission. Although learning-based lossless compressors achieve strong results, most of them are designed for a single modality, leading to redundant compressor deployments in multi-modal settings. Designing a unified multi-modal compressor is critical yet challenging, as different data types vary largely in format, dimension, and statistics. Multi-modal large language models offer a promising resolution but remain too complex for practical use. Thus, we propose OmniZip, a unified and lightweight lossless compressor for multi-modal data (like image, text, speech, tactile, database, and gene sequence). Built on a lightweight backbone, OmniZip incorporates three key components to enable efficient multi-modal lossless compression: a modality-unified tokenizer that reversibly transforms diverse data into tokens, a modality-routing context learning mechanism that enables flexible multi-modal context modeling, and a modality-routing feedforward design that further enhances the model's nonlinear representation flexibility. A reparameterization training strategy is used to enhance model capacity. OmniZip outperforms or matches other state-of-the-art compressors on multiple modalities, achieving 42%, 57%, 62% and 42%, 53% higher compression efficiency than gzip on CLIC-M, TouchandGo, enwik9, LibriSpeech, and WikiSQL datasets, respectively. It also supports near real-time inference on resource-constrained edge devices, reaching about 1MB/s on MacBook CPUs and iPhone NPUs. Our code is released at https:// github.com/adminasmi/OmniZip-CVPR2026.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Language Modeling Is CompressionGrégoire Delétang, Anian Ruoss, Paul-Ambroise Duquenne, Elliot Catt 等ICLR 2024 · 被引用 243 次
- iFlow: Numerically Invertible Flows for Efficient Lossless Compression via a Uniform CoderShifeng Zhang, Ning Kang, Tom Ryder, Zhenguo LiNeurIPS 2021 · 被引用 47 次
- Large Language Models for Lossless Image Compression: Next-Pixel Prediction in Language Space is All You NeedKecheng Chen, Pingping Zhang, Hui Liu, Jie Liu 等NeurIPS 2025 · 被引用 14 次
- MSDZip: Universal Lossless Compression for Multi-source Data via Stepwise-parallel and Learning-based PredictionHuidong Ma, Hui Sun, Liping Yi, Yanfeng Ding 等WWW 2025 · 被引用 9 次
- L3TC: Leveraging RWKV for Learned Lossless Low-Complexity Text CompressionJunxuan Zhang, Zhengxue Cheng, Yan Zhao, Shihao Wang 等AAAI 2025 · 被引用 8 次
相关 Paper
- OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language ModelsKeda Tao, Kele Shao, Bohan Yu, Weiqiang Wang 等CVPR 2026 · 被引用 32 次
- OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language ModelsYue Ding, Yiyan Ji, Jungang Li, Xuyang Liu 等ICML 2026 · 被引用 22 次
- Compression via Pre-trained Transformers: A Study on Byte-Level Multimodal DataDavid Heurtel-Depeiges, Anian Ruoss, Joel Veness, Tim GeneweinICML 2025
- UniCompress: Token Compression for Unified Vision-Language Understanding and GenerationZiyao Wang, Chen Chen, Jingtao Li, Weiming Zhuang 等CVPR 2026 · 被引用 1 次
- MoE-LC: General-Purpose Lossless Compression for Multi-modal Data via Entropy-Aware Multi-ExpertsZeyi Lu, Xiaoxiao Ma, Yujun Huang, Minxiao Chen 等WWW 2026
