Score: 0

CorGi: Contribution-Guided Block-Wise Interval Caching for Training-Free Acceleration of Diffusion Transformers

Published: December 30, 2025 | arXiv ID: 2512.24195v1

By: Yonglak Son , Suhyeok Kim , Seungryong Kim and more

Diffusion transformer (DiT) achieves remarkable performance in visual generation, but its iterative denoising process combined with larger capacity leads to a high inference cost. Recent works have demonstrated that the iterative denoising process of DiT models involves substantial redundant computation across steps. To effectively reduce the redundant computation in DiT, we propose CorGi (Contribution-Guided Block-Wise Interval Caching), training-free DiT inference acceleration framework that selectively reuses the outputs of transformer blocks in DiT across denoising steps. CorGi caches low-contribution blocks and reuses them in later steps within each interval to reduce redundant computation while preserving generation quality. For text-to-image tasks, we further propose CorGi+, which leverages per-block cross-attention maps to identify salient tokens and applies partial attention updates to protect important object details. Evaluation on the state-of-the-art DiT models demonstrates that CorGi and CorGi+ achieve up to 2.0x speedup on average, while preserving high generation quality.

ProCache: Constraint-Aware Feature Caching with Selective Computation for Diffusion Transformer Acceleration

CV and Pattern Recognition

Makes AI image creation much faster.

19 Dec 2025 1

89%

MixCache: Mixture-of-Cache for Video Diffusion Transformer Acceleration

Graphics

Makes videos create faster without losing quality.

18 Aug 2025 0

89%

GalaxyDiT: Efficient Video Generation with Guidance Alignment and Adaptive Proxy in Diffusion Transformers

CV and Pattern Recognition

Makes videos faster without losing quality.

3 Dec 2025 3

View PDF Login to Bookmark

CorGi: Contribution-Guided Block-Wise Interval Caching for Training-Free Acceleration of Diffusion Transformers

Technical Abstract

ProCache: Constraint-Aware Feature Caching with Selective Computation for Diffusion Transformer Acceleration

MixCache: Mixture-of-Cache for Video Diffusion Transformer Acceleration

GalaxyDiT: Efficient Video Generation with Guidance Alignment and Adaptive Proxy in Diffusion Transformers