SmoothCache uses calibration-measured layer error thresholds to skip redundant attention and feed-forward computations in Diffusion Transformers, achieving 8-71% speedup across image, video, and audio tasks.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
SmoothCache: A Universal Inference Acceleration Technique for Diffusion Transformers
SmoothCache uses calibration-measured layer error thresholds to skip redundant attention and feed-forward computations in Diffusion Transformers, achieving 8-71% speedup across image, video, and audio tasks.