SCOPE accelerates video diffusion transformers about 2x by scoring keys through 3D-RoPE subspace clusters and adaptively setting per-head Top-k counts, matching dense-attention fidelity closely.
Generalized neighborhood attention: Multi-dimensional sparse attention at the speed of light
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
SCOPE: Subspace Clustering with Online Per-Head Top-K Estimation for Sparse Video Attention
SCOPE accelerates video diffusion transformers about 2x by scoring keys through 3D-RoPE subspace clusters and adaptively setting per-head Top-k counts, matching dense-attention fidelity closely.