ASDSV: Multimodal Generation Made Efficient with Approximate Speculative Diffusion and Speculative Verification
Kaijun Zhou, Xingyu Yan, Xingda Wei, Xijun Li, Jinyu Gu
摘要
Diffusion in transformer is central to advances in high-quality multimodal generation but suffer from high inference latency due to their iterative nature. Inspired by speculative decoding’s success in accelerating large language models, we pro-pose Approximate Speculative Diffusion with Speculative Verification (ASDSV) , a novel method to enhance the efficiency of diffusion models. Adapting speculative execution to diffusion processes presents unique challenges. First, the substantial computational cost of verifying numerous speculative steps for continuous, high-dimensional outputs makes traditional full verification pro-hibitively expensive. Second, determining the optimal number of speculative steps K involves a trade-off between potential acceleration and verification success rates. To address these, ASDSV introduces two key innovations: 1) A speculative verification technique, which leverages the observed temporal correlation between draft and target model outputs, efficiently validates K speculative steps by only checking the alignment of the initial and final states, significantly reducing verification
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper24
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari 等ICML 2024 · 被引用 3,620 次
相关 Paper
- DFlash: Block Diffusion for Flash Speculative DecodingJian Chen, Yesheng Liang, Zhijian LiuICML 2026
- SpeCa: Accelerating Diffusion Transformers with Speculative Feature CachingJiacheng Liu, Chang Zou, Yuanhuiyi Lyu, Fei Ren 等ACM MM 2025 · 被引用 3 次
- Dynamic-Width Speculative Beam Decoding for LLM InferenceZongyue Qin, Zifan He, Neha Prakriya, Jason Cong 等AAAI 2025 · 被引用 10 次
- AASD: Accelerate Inference by Aligning Speculative Decoding in Multimodal Large Language ModelsChaoqun Yang, Ran Chen, Muyang Zhang, Weiguang Pang 等DAC 2025 · 被引用 1 次
- Traversal Verification for Speculative Tree DecodingYepeng Weng, Qiao Hu, Xujie Chen, Li Liu 等NeurIPS 2025 · 被引用 11 次
