Revisiting Redundancy in Diffusion Transformers: A Temporal-Spatial Joint Caching Strategy for Efficient Sampling
Chenxi Du, Yongheng Deng, Ju Ren, Yaoxue Zhang
摘要
Diffusion Transformers (DiTs) achieve impressive generative performance but suffer from significant inference latency. Feature caching–based acceleration methods reduce total computation by reusing results from earlier timesteps, but they largely ignore that temporal redundancy is dynamic and inconsistent across timesteps. Our analysis reveals this variability. More crucially, we identify a previously underexplored form of efficiency, namely spatial redundancy, characterized by high similarity between adjacent transformer blocks within the same timestep. Motivated by this dual-dimensional redundancy, we propose Temporal-Spatial Joint Cache, a training-free inference acceleration strategy that dynamically determines optimal reuse operations across temporal and spatial dimensions. Our approach features a redundancy-guided operation selector that estimates local feature stability using second-order divided differences, enabling fine-grained decisions between full computation, temporal cache, and spatial cache. Furthermore, we use interpolation-based feature prediction to capture local feature evolution for more accurate reuse. In addition, we propose a bounded cache distance control mechanism to mitigate error accumulation from excessive reuse. Together, these components allow our method to deliver substantial inference speedups without retraining or compromising generation fidelity, offering a new perspective on efficiency in diffusion transformer inference.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- BWCache: Accelerating Video Diffusion Transformers through Block-Wise CachingHanshuai Cui, Zhiqing Tang, Zhifei Xu, Zhi Yao 等ICLR 2026 · 被引用 11 次
- Sortblock: Similarity-Aware Feature Reuse for Diffusion ModelHanqi Chen, Xu Zhang, Xiaoliu Guan, Lielin Jiang 等AAAI 2026
- ProCache: Constraint-Aware Feature Caching with Selective Computation for Diffusion Transformer AccelerationFanpu Cao, Yaofo Chen, Zeng You, Wei LuoAAAI 2026
- BlockDance: Reuse Structurally Similar Spatio-Temporal Features to Accelerate Diffusion TransformersHui Zhang, Tingwei Gao, Jie Shao, Zuxuan WuCVPR 2025
- FasterCache: Training-Free Video Diffusion Model Acceleration with High QualityZhengyao Lv, Chenyang Si, Junhao Song, Zhenyu Yang 等ICLR 2025
