ERTACache: Error Rectification and Timesteps Adjustment for Efficient Diffusion
Xurui Peng, Chenqian Yan, Hong Liu, Rui Ma, Fangmin Chen, Xing Wang, Zhihua Wu, Songwei Liu, Mingbao Lin
Abstract
Diffusion models suffer from substantial computational overhead due to their inherently iterative inference process. While feature caching offers a promising acceleration strategy by reusing intermediate outputs across timesteps, naive reuse often incurs noticeable quality degradation. In this work, we formally analyze the cumulative error introduced by caching and decompose it into two principal components: feature shift error, caused by inaccuracies in cached outputs, and step amplification error, which arises from error propagation under fixed timestep schedules. To address these issues, we propose ERTACache, a principled caching framework that jointly rectifies both error types. Our method employs an offline residual profiling stage to identify reusable steps, dynamically adjusts integration intervals via a trajectory-aware correction coefficient, and analytically approximates cache-induced errors through a closed-form residual linearization model. Together, these components enable accurate and efficient sampling under aggressive cache reuse. Extensive experiments across standard image and video generation benchmarks show that ERTACache achieves up to 2x inference speedup while consistently preserving or even improving visual quality. Notably, on the state-of-the-art Wan2.1 video diffusion model, ERTACache delivers 2x acceleration with minimal VBench degradation, effectively maintaining baseline fidelity while significantly improving efficiency. The code is available at https://github.com/bytedance/ERTACache.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f43b2364-d6a9-4d21-bc51-1d9b214b7929Cited by top-tier papers3
- EigenCache: Rethinking Diffusion Acceleration as Covariance-Optimal Forecasting and Submodular Information AllocationChenyang Xu, Dezhen Wang, Lin Chen, Kepeng Lin et al.ICML 2026
- Motion-Aware Caching for Efficient Autoregressive Video GenerationJing Xu, Yuexiao Ma, Xuzhe Zheng, WANG et al.ICML 2026
- S2O: Early Stopping for Sparse Attention via Online PermutationYu Zhang, Songwei Liu, Chenqian Yan, Sheng Lin et al.ACL 2026
Builds on20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari et al.ICML 2024 · 3,620 citations
Related papers
- SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware CachingYasaman Haghighi, Alexandre AlahiCVPR 2026 · 5 citations
- D2Cache: Second-Order Delta Caching for Higher Video Diffusion AccelerationEnhuai Liu, Yunke Wang, Changming Sun, Chang XuCVPR 2026
- Accelerating Autoregressive Video Diffusion via History-Guided Cache and Residual CorrectionKepan Nan, Wangbo Zhao, Penghao Zhou, Jun Li et al.CVPR 2026
- DiCache: Let Diffusion Model Determine Its Own CacheJiazi Bu, Pengyang Ling, Yujie Zhou, Yibin Wang et al.ICLR 2026 · 36 citations
- Model Reveals What to Cache: Profiling-Based Feature Reuse for Video Diffusion ModelsXuran Ma, Yexin Liu, Yaofu Liu, Xianfeng Wu et al.ICCV 2025 · 16 citations
