OSCAR: One-Step Diffusion Codec Across Multiple Bit-rates
Jinpei Guo, Yifei Ji, Zheng Chen, Kai Liu, Ming Liu, Wang Rao, Wenbo Li, Yong Guo, Yulun Zhang
摘要
Pretrained latent diffusion models have shown strong potential for lossy image compression, owing to their powerful generative priors. Most existing diffusion-based methods reconstruct images by iteratively denoising from random noise, guided by compressed latent representations. While these approaches have achieved high reconstruction quality, their multi-step sampling process incurs substantial computational overhead. Moreover, they typically require training separate models for different compression bit-rates, leading to significant training and storage costs. To address these challenges, we propose a one-step diffusion codec across multiple bit-rates. termed OSCAR. Specifically, our method views compressed latents as noisy variants of the original latents, where the level of distortion depends on the bit-rate. This perspective allows them to be modeled as intermediate states along a diffusion trajectory. By establishing a mapping from the compression bit-rate to a pseudo diffusion timestep, we condition a single generative model to support reconstructions at multiple bit-rates. Meanwhile, we argue that the compressed latents retain rich structural information, thereby making one-step denoising feasible. Thus, OSCAR replaces iterative sampling with a single denoising pass, significantly improving inference efficiency. Extensive experiments demonstrate that OSCAR achieves superior performance in both quantitative and visual quality metrics. The code and models are available at https://github.com/jp-guo/OSCAR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Ultra-Low Bitrate Perceptual Image Compression with Shallow EncoderTianyu Zhang, Dong Liu, Chang Wen ChenCVPR 2026 · 被引用 5 次
- Compression-Aware One-Step Diffusion Model for JPEG Artifact RemovalJinpei Guo, Zheng Chen, Wenbo Li, Yong Guo 等ICCV 2025 · 被引用 5 次
- Steering One-Step Diffusion Model with Fidelity-Rich Decoder for Fast Image CompressionZheng Chen, Mingde Zhou, Jinpei Guo, Jiale Yuan 等AAAI 2026 · 被引用 1 次
- Efficient Learned Image Compression without Entropy CodingHao Cao, Wenqi Guo, Zhijin Qin, Jungong HanICML 2026
它引用的顶会 Paper27
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 被引用 11,724 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
相关 Paper
- StableCodec: Taming One-Step Diffusion for Extreme Image CompressionTianyu Zhang, Xin Luo, Li Li, Dong LiuICCV 2025 · 被引用 7 次
- One-Step Diffusion-Based Image Compression with Semantic DistillationNaifu Xue, Zhaoyang Jia, Jiahao Li, Bin Li 等NeurIPS 2025 · 被引用 28 次
- One-Step Effective Diffusion Network for Real-World Image Super-ResolutionRongyuan Wu, Lingchen Sun, Zhiyuan Ma, Lei ZhangNeurIPS 2024 · 被引用 319 次
- Generative Latent Diffusion for Efficient Spatiotemporal Data ReductionXiao Li, Liangji Zhu, Anand Rangarajan, Sanjay RankaSC 2025 · 被引用 1 次
- DOVE: Efficient One-Step Diffusion Model for Real-World Video Super-ResolutionZheng Chen, Zichen Zou, Kewei Zhang, Xiongfei Su 等NeurIPS 2025 · 被引用 31 次
