BiMaCoSR: Binary One-Step Diffusion Model Leveraging Flexible Matrix Compression for Real Super-Resolution
Kai Liu, Kaicheng Yang, Zheng Chen, Zhiteng Li, Yong Guo, Wenbo Li, Linghe Kong, Yulun Zhang
摘要
While super-resolution (SR) methods based on diffusion models (DM) have demonstrated inspiring performance, their deployment is impeded due to the heavy request of memory and computation. Recent researchers apply two kinds of methods to compress or fasten the DM. One is to compress the DM into 1-bit, aka binarization, alleviating the storage and computation pressure. The other distills the multi-step DM into only one step, significantly speeding up inference process. Nonetheless, it remains impossible to deploy DM to resource-limited edge devices. To address this problem, we propose BiMaCoSR, which combines binarization and one-step distillation to obtain extreme compression and acceleration. To prevent the catastrophic collapse of the model caused by binarization, we propose sparse matrix branch (SMB) and low rank matrix branch (LRMB). Both auxiliary branches pass the full-precision (FP) information but in different ways. SMB absorbs the extreme values and its output is high rank, carrying abundant FP information. Whereas, the design of LRMB is inspired by LoRA and is initialized with the top r SVD components, outputting low rank representation. The computation and storage overhead of our proposed branches can be safely ignored. Comprehensive comparison experiments are conducted to exhibit BiMaCoSR outperforms current stateof-the-art binarization methods and gains competitive performance compared with FP one-step model. BiMaCoSR achieves a 23.8× compression ratio and a 27.4× speedup ratio compared to FP counterpart. Our code and model are available at https://github.com/Kai-Liu001/BiMaCoSR .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- RobuQ: Pushing DiTs to W1.58A2 via Robust Activation QuantizationKaicheng Yang, Xun Zhang, Haotong Qin, Yucheng Lin 等ICML 2026 · 被引用 5 次
- Q-DiT4SR: Exploration of Detail-Preserving Diffusion Transformer Quantization for Real-World Image Super-ResolutionXun Zhang, Kaicheng Yang, Hongliang Lu, Haotong Qin 等ICML 2026 · 被引用 2 次
- InfVSR: Toward Consistency-Driven Streaming Generative Video Super-ResolutionZiqing Zhang, Kai Liu, Zheng Chen, Xi Li 等ICML 2026
- TWLA: Achieving Ternary Weights and Low-Bit Activations for LLMs via Post-Training QuantizationZhixiong Zhao, Zukang Xu, Zhixuan Chen, Xing Hu 等ICML 2026
它引用的顶会 Paper19
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou 等CVPR 2022 · 被引用 1,970 次
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 被引用 1,208 次
- Designing a Practical Degradation Model for Deep Blind Image Super-ResolutionKai Zhang, Jingyun Liang, Luc Van Gool, Radu TimofteICCV 2021 · 被引用 898 次
- Toward Real-World Single Image Super-Resolution: A New Benchmark and a New ModelJianrui Cai, Hui Zeng, Hongwei Yong, Zisheng Cao 等ICCV 2019 · 被引用 713 次
相关 Paper
- Binarized Diffusion Model for Image Super-ResolutionZheng Chen, Haotong Qin, Yong Guo, Xiongfei Su 等NeurIPS 2024 · 被引用 36 次
- Restabilizing Diffusion Models with Predictive Noise Fusion Strategy for Image Super-ResolutionLuoqian Jiang, Yong Guo, Bingna Xu, Haolin Pan 等AAAI 2025
- BiDM: Pushing the Limit of Quantization for Diffusion ModelsXingyu Zheng, Xianglong Liu, Yichen Bian, Xudong Ma 等NeurIPS 2024 · 被引用 12 次
- BiMatting: Efficient Video Matting via BinarizationHaotong Qin, Lei Ke, Xudong Ma, Martin Danelljan 等NeurIPS 2023 · 被引用 28 次
- BinaryDM: Accurate Weight Binarization for Efficient Diffusion ModelsXingyu Zheng, Xianglong Liu, Haotong Qin, Xudong Ma 等ICLR 2025
