Pith. sign in

Paper Citation Record · LEDGER

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

As of 21 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 72 inbound Pith citation observations for arXiv:2502.01776.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01776 v2

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T14:35:10.707673Z

measured 143 of 143 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 72 of 72 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:38:39.235933Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:49:44.815320Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact2
  • verified fuzzy22
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 64c08113-569a-41c4-8077-821c6f92b04d · outbound

This paper cites Vivit: A video vision transformer.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Vivit: A video vision transformer

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:12.052716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.400597Z digest=sha256:f2a0d3db3f4cf548c3e71bd233e12104b9a9aa834b654a3b3dd60ba9c5082888

Observation 1239322d-3d74-46fa-9bf6-cfec7a595b5e · outbound

This paper cites Efficientvit: Lightweight multi-scale attention for high-resolution dense prediction.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Efficientvit: Lightweight multi-scale attention for high-resolution dense prediction

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:12.038205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.405349Z digest=sha256:ad7c3d86dca8ec3be1dd26fc96b3dc3d7bc5e104ce92e8fa6b000d58e4da8041

Observation 653a7eef-abd8-48d7-b1f9-8ace9835ffeb · outbound

This paper cites Condition-aware neural network for controlled image generation.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Condition-aware neural network for controlled image generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:12.022339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.410009Z digest=sha256:6ce9aa8b45d294843bf9236e08d6d2c52d86425646a3e5404ca9c0c3a81d0e42

Observation fbcc3c14-5c10-451d-9f2d-1e0a95cb4718 · outbound

This paper cites Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.414430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.414430Z digest=sha256:caa92f8264610bb73e6777ec3ff60902e5da9540c747d8bf7d78da11ee5ad45b

Observation 0749b57f-061e-4a4f-a8bb-befaf8380df3 · outbound

This paper cites Pixart-sigma: Weak-to-strong training of diffusion transformer for 4k text-to-image generation.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Pixart-sigma: Weak-to-strong training of diffusion transformer for 4k text-to-image generation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:12.008132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.419083Z digest=sha256:6c4b0091cb9c4fe56128548295614a3d10cedfc5fcbc44a6c887517cf42b6511

Observation db95ab49-8eaa-4020-b584-c102977fd18f · outbound

This paper cites $\Delta$-DiT: A Training-Free Acceleration Method Tailored for Diffusion Transformers.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity $\Delta$-DiT: A Training-Free Acceleration Method Tailored for Diffusion Transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.423423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.423423Z digest=sha256:824f3a3cff8427d47b3c129a292ee5b2782bde59d76300d43831fb990aac5dcd

Observation fbd9c6d6-09b6-40e7-abd3-55df14e4e780 · outbound

This paper cites Rethinking Attention with Performers.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Rethinking Attention with Performers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.428626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.428626Z digest=sha256:e71c82e1cd1a00778581b7f3af991d2076e2825fe6839cc7c5baf8c97289dcf3

Observation bd97d16f-5280-448a-95c4-483678caf532 · outbound

This paper cites Learning fast algorithms for linear transforms using butterfly factorizations.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Learning fast algorithms for linear transforms using butterfly factorizations

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.993960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.433069Z digest=sha256:a7d71106249e7ee8094bb90486fe268b9673a185dac194890b41d2de10a0c2ea

Observation 56cc4516-3c64-47ca-b8ab-55e24a3e73b7 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.437142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.437142Z digest=sha256:0a2e95885e10a78ed9acf02ce796bd1918cdc54f44c5d72fa5ec8f73bfb69882

Observation 7940b615-3791-4aef-a09e-2d3274149306 · outbound

This paper cites J., and Zhang, X.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity J., and Zhang, X

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.979946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.441842Z digest=sha256:324b2fb2d3d7835701d79e04c2d8c8eadfa6230ae185dbcb366b9b5dbaaef5bf

Observation be0ce97c-2c41-491f-87e7-df2356fc4d3c · outbound

This paper cites Animatediff: Animate your personalized text-to-image diffusion models without specific tuning.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Animatediff: Animate your personalized text-to-image diffusion models without specific tuning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.965544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.446267Z digest=sha256:db4c228847dfddef4f3d8be3341167b9d7a3520e074766690e9b44a262ab1b92

Observation 6afc8d7c-dbe0-460f-a695-dfda2b3c5830 · outbound

This paper cites LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.450502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.450502Z digest=sha256:1cedf6b33ae8e68d16bc15fb9e5c72d7f17e46752bf8c6eedde0a85fcb4aaff2

Observation dc75dd23-953d-4e9d-9125-4ea62f6be259 · outbound

This paper cites Denoising diffusion probabilistic models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Denoising diffusion probabilistic models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.455002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.455002Z digest=sha256:401d901ca08d3b2286650a7098939769374fd794a7221df26b60879d7ed5aa6a

Observation 1408dd31-3412-4fdc-9069-4c9d3ddf440a · outbound

This paper cites Cogvideo: Large-scale pretraining for text-to-video generation via transformers.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Cogvideo: Large-scale pretraining for text-to-video generation via transformers

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.942103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.459236Z digest=sha256:93939a376daae03eeb5d928014647980293d2b120a8c6e9c06f1aaed09f68ca9

Observation c55e4273-df53-4f7a-abfd-85940d5bc0a3 · outbound

This paper cites and Ziou, D.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity and Ziou, D

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.463385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.463385Z digest=sha256:ed58187516a3ee46ff8e9b961d8a3a5f956f65b4b99639814937764ec32db08a

Observation 3ed40a54-c4f0-4168-9c20-bdaac8171774 · outbound

This paper cites VBench: Comprehensive Benchmark Suite for Video Generative Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.467966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.467966Z digest=sha256:802ef943753b4774ee4c55a613eb64e4a5889a805a4072665afae13bdab2331f

Observation cd041ad0-0ee2-45fc-9301-b0f8a83ec79d · outbound

This paper cites MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.472320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.472320Z digest=sha256:ee91df116ef58486ce8ba832c4c3b10283933d0f9e2d9d8945f292051e4b743a

Observation 6846e508-5e6c-4ea9-ade7-cb896d39a758 · outbound

This paper cites Transformers are rnns: Fast autoregressive transformers with linear attention.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Transformers are rnns: Fast autoregressive transformers with linear attention

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.476898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.476898Z digest=sha256:4bc9d531d43b397a5fdd08a6413c9689ebf293c63412a8808935316864e17e72

Observation 9553d3cd-8ef8-45cf-b906-7b61c4e7a636 · outbound

This paper cites StreamDiffusion: A Pipeline-level Solution for Real-time Interactive Generation.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity StreamDiffusion: A Pipeline-level Solution for Real-time Interactive Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.481230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.481230Z digest=sha256:eac4ddf647f8c6d5b94231d9045c4b6705f1f3dd263560139a133923926c1f41

Observation 6a5a9613-e8a3-40f8-add5-05f10e4c8465 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.485622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.485622Z digest=sha256:6865104fdbb0cd906012b383001fd01332689dc82e016e13a6cf75e3eb88e030

Observation fe4b1987-9a3c-4954-894f-3be11272fa22 · outbound

This paper cites Distrifusion: Distributed parallel inference for high-resolution diffusion models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Distrifusion: Distributed parallel inference for high-resolution diffusion models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.918311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.489998Z digest=sha256:1363544dcb100ce92556dfa646ed67bd6a2643547231be0bd3c7944a01cb03aa

Observation c4ffc882-7c80-4f7d-95c5-d2dc065117b4 · outbound

This paper cites Svdquant: Absorbing outliers by low-rank components for 4-bit diffusion models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Svdquant: Absorbing outliers by low-rank components for 4-bit diffusion models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.904119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.494219Z digest=sha256:5b56c82bcf60a76f7afe77dd2ac9776bc35505870fa9fea25c4739dbe2b151b0

Observation f90ad5c5-ff4e-46e9-8ec6-50d998706563 · outbound

This paper cites Q-diffusion: Quantizing diffusion models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Q-diffusion: Quantizing diffusion models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.890408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.498423Z digest=sha256:dd4a8851971f537f74ceed3fb0a0a349c67e6ffdf1256992608a8387b6f18a45

Observation 2676df55-5473-431b-86d0-a84c9306722c · outbound

This paper cites MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.502622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.502622Z digest=sha256:a93fef2259def678379b6761bf446a97b4ebdf544159cc064021cb839d8be801

Observation c2213898-ce79-40bb-bfaf-eaf5baf0b235 · outbound

This paper cites Looking Backward: Streaming Video-to-Video Translation with Feature Banks.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Looking Backward: Streaming Video-to-Video Translation with Feature Banks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.507059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.507059Z digest=sha256:a8fab8c530eceb17fa6d8997916a3ebb7a2d6079957a467bdf437f5da8e75e6c

Observation 63e184d8-eab0-4bb3-99e9-24c9409942c3 · outbound

This paper cites Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.511735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.511735Z digest=sha256:8f907ea8072ce22aa4d4d10165f41f3483639710d8ed7254005445e59a170d26

Observation e184dcdd-2278-4697-a129-9408c1a4d8e9 · outbound

This paper cites World model on million-length video and language with ringattention.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity World model on million-length video and language with ringattention

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.876738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.516381Z digest=sha256:9ab6c5740c8f7b162405acce245d75a8a52f917fcf7daf9d33f870b3f6b7417e

Observation d6041352-3b72-4494-8528-d46cea195395 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.520415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.520415Z digest=sha256:c18d6e160f1c662586efa4c9cf5ed9c19197b649fbb03f01b7236b742b63a3d0

Observation 942776ef-3b45-44fd-86e9-69083e2659d4 · outbound

This paper cites Instaflow: One step is enough for high-quality diffusion-based text-to-image generation.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Instaflow: One step is enough for high-quality diffusion-based text-to-image generation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.861522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.524854Z digest=sha256:3aded1d615877cf1384723cee301d1ddfc111e79c0fc695b874136b863974268

Observation c1227991-1569-420b-97e9-bd172a54664d · outbound

This paper cites Scissorhands: Exploiting the persistence of importance hypothesis for llm kv cache compression at test time.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Scissorhands: Exploiting the persistence of importance hypothesis for llm kv cache compression at test time

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.845703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.528939Z digest=sha256:d569f9466dc66900332a472a78c9358bb1ae492e99df6a0b35d9a028fc20f179

Observation 5b32ecea-b2ba-45be-a2a9-cef03d838ebe · outbound

This paper cites Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.831483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.532860Z digest=sha256:2320b62f4b1a62170ef1cb5c99ac920bf8e62cd16008df4ca4d0ffc08fb5c7b2

Observation ad298745-1c44-4ab0-91d5-71e2e9ec3f70 · outbound

This paper cites DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.536888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.536888Z digest=sha256:cccd9f240d706edc3a3ef17d47a0c86428e827223750a7bbf0f011eba756a261

Observation 918fab8a-0e0d-47fb-a163-864dfb3874a6 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.541246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.541246Z digest=sha256:42e9a083a539bed35ce9687455ef462956f107b84e1109a2916d7cd74f1e94f9

Observation 6e8312db-b56e-4ebd-9aa5-4c8384775ef5 · outbound

This paper cites an unresolved cited work.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-09T14:35:11.817027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.545495Z digest=sha256:6b164885fbcae2be76f8e92dc87f35966357a734a150b1da8f3b0ff09625445f

Observation b4a7a82d-6f0f-45e2-a0f8-77339c8fa532 · outbound

This paper cites Deepcache: Accelerating diffusion models for free.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Deepcache: Accelerating diffusion models for free

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.802914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.549647Z digest=sha256:fdd2b7c122c4a125f72da8e80950f652e6b06735dce884a6735f9e4c225d5425

Observation 185ba687-b17e-4c2b-881f-98eee694e6df · outbound

This paper cites SDE dit: Guided image synthesis and editing with stochastic differential equations.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity SDE dit: Guided image synthesis and editing with stochastic differential equations

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.788588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.553693Z digest=sha256:74c314ddb946dd6205b1e7b3745f30ab8c36eae43151c1298fbe4f158d9bc8b0

Observation 92a21a64-a7ff-49d6-8bfb-4e1f3bbc59a5 · outbound

This paper cites and Xie, S.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity and Xie, S

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.557648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.557648Z digest=sha256:279131448f204df80010f9231f08b93a74a0209d172843b838a0ea07f1c4532e

Observation bcc70621-8882-4b7b-841e-bdbfb7b75196 · outbound

This paper cites LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.561764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.561764Z digest=sha256:88ea1ca98d8ce3c41eb7ed02c619e231d32c25ddee0192f63a5dd047866e0047

Observation 770b6f5a-c98b-4b0e-a26f-ee7c029ed606 · outbound

This paper cites Denoising diffusion implicit models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Denoising diffusion implicit models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.766013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.566100Z digest=sha256:df9dcb66e78976bb85f98d7bf7a75ee61a0d44cc3b9510ea140608240e56bb03

Observation 0c3c93b7-a4c6-4908-9b66-002c2478a4ee · outbound

This paper cites and Ermon, S.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity and Ermon, S

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.570139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.570139Z digest=sha256:6675f64640bee1cfbbc12f066f79d43a70022b27e32c550096a53aaeb185167e

Observation c7f5da32-d8c6-4dd4-b4a6-656a836f83ed · outbound

This paper cites Consistency models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Consistency models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.574246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.574246Z digest=sha256:c0715911cde0e2095ab86db9200520369eaccad95ed3e8dd336ef6c41fa5c73c

Observation 6a3f4b03-3f62-4b49-a654-dd4d3dab1e07 · outbound

This paper cites AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.578399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.578399Z digest=sha256:91a95434c77f587f4901208473fd1205424c087bbc6f4badb567119dcf088a9b

Observation 3c96c00c-118e-4a65-8471-bbc61162ec7e · outbound

This paper cites Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.583214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.583214Z digest=sha256:6ddda2c2f005a60a1e1283f9da6945d7bc1779623fc2280472045ad9c2bfee3c

Observation c63b42de-c093-4d62-a185-73c12c6a6cd8 · outbound

This paper cites an unresolved cited work.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.587706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.587706Z digest=sha256:1a70b644e01ea4d0a38dca9d2df858f769ac82c77218ff1a675bcedcf9798281

Observation c236b263-07e4-4530-9cc3-613dd798d544 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Wan: Open and Advanced Large-Scale Video Generative Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.591901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.591901Z digest=sha256:6089a3257efe905a02ee2d956bffde1dc50bdd344fcae8fc3cb84e2b662eeaa5

Observation 6b69baac-df58-4fd1-b00e-d9a47541b824 · outbound

This paper cites Linformer: Self-Attention with Linear Complexity.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Linformer: Self-Attention with Linear Complexity

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.596281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.596281Z digest=sha256:480b43d4a63673e15d6cdd3a21b25b74b74ebcad479c717d26f751dec0175ca7

Observation 037a8f4d-a0a2-4430-9819-cc354bc684e0 · outbound

This paper cites DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.604722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.604722Z digest=sha256:cbf59ee8bb311c9609e8b3bb5fad967e9510e28c551dc6bed40a26049871e3ad

Observation a78ba5c5-3977-4caf-8009-47cfb7fdb361 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Efficient Streaming Language Models with Attention Sinks

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.608973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.608973Z digest=sha256:b4b4106679e706b9049a47929acbc9cc91955dbb186ad1ef0a0ec10c9476502e

Observation 3935e04b-58fe-40c2-86db-b89a76aca4e2 · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.612976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.612976Z digest=sha256:175b0c7ce10f22168bd265bea1314f4be0d63db3b83e63d4ad24c186771a9be1

Observation 6312eccf-126e-4ff2-a4b7-7eb7dc470a91 · outbound

This paper cites XAttention: Block Sparse Attention with Antidiagonal Scoring.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity XAttention: Block Sparse Attention with Antidiagonal Scoring

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.617359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.617359Z digest=sha256:404d670b0f998992a859bd132c309a12178a536c08fe51cdb80b3a3b55839c9b

Observation 1c1f235a-f392-4caa-9449-c6603fb66594 · outbound

This paper cites TidalDecode: Fast and Accurate LLM Decoding with Position Persistent Sparse Attention.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity TidalDecode: Fast and Accurate LLM Decoding with Position Persistent Sparse Attention

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.621827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.621827Z digest=sha256:50030083bf467169e249e67f0c9cc53ef065c3a486bb9f6b71ffa430981bcf0a

Observation 3a65e566-1065-4199-b68d-008a03315795 · outbound

This paper cites Post-Training Sparse Attention with Double Sparsity.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Post-Training Sparse Attention with Double Sparsity

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.626233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.626233Z digest=sha256:80eb2dea77269fed9b49439183bc9226efa1527fe9e9f633e31f694a5245c692

Observation 822b0ac8-b241-4f8e-8be1-4ac1770026ea · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.630494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.630494Z digest=sha256:fd56c00b00708f20b85b038af9020232f63bc47862baf669d41cb3105ba11f1b

Observation dcaabfb5-ae5b-4fa3-a16a-f4a9a216905c · outbound

This paper cites SparseTIR: Composable Abstractions for Sparse Compilation in Deep Learning.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity SparseTIR: Composable Abstractions for Sparse Compilation in Deep Learning

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-09T14:35:11.268172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.634908Z digest=sha256:032aa74146e9bec1fbd8a5221a1c9d4bd711550f1fd585b81c2ddaf02d9a65e7

Observation 5ad4b3cb-c827-4629-8f5c-1b2efe4464ef · outbound

This paper cites FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.639239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.639239Z digest=sha256:781814be4b60ced7cb65c77783613b4d9847abdceae997ff85159ae2943ba553

Observation 229bafaa-2b98-4c71-8bb5-c88cbc50d78a · outbound

This paper cites Improved Distribution Matching Distillation for Fast Image Synthesis.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Improved Distribution Matching Distillation for Fast Image Synthesis

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.643588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.643588Z digest=sha256:925d278c0494081233c88d1d33779ffce9466e9d766dcf0c6a13f5ee705ff937

Observation 8b8be848-fd81-4c7f-89ef-d97f3b89d370 · outbound

This paper cites T., and Park, T.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity T., and Park, T

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.647847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.647847Z digest=sha256:b1c216db0eb38e456d6b7624196447dbfcb685976c75a796901501b0a2fa2097

Observation e26ed329-929d-4fb4-adf1-1ed5d9b74290 · outbound

This paper cites Metaformer is actually what you need for vision.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Metaformer is actually what you need for vision

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.652411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.652411Z digest=sha256:3630909f1bb024dc2db67a87db3bb61cdf1cf4dbf914c219d02a4e922ab14c31

Observation 3fc96d14-e17a-4868-956e-0a68bd083f52 · outbound

This paper cites DiTFastAttn: Attention Compression for Diffusion Transformer Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity DiTFastAttn: Attention Compression for Diffusion Transformer Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.656633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.656633Z digest=sha256:d7e93ec837c04f2141f6d1ceb55a037584f20f8f7d9e1f6e26b750d463e55adf

Observation 3185ecba-cee6-43f6-80c2-7203dd6df08d · outbound

This paper cites P., Jampani, V., Sun, D., and Yang, M.-H.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity P., Jampani, V., Sun, D., and Yang, M.-H

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.708764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.660922Z digest=sha256:e2963cb6d462007145af3f254c47529abce2d01658491b3ad32d615398c5592f

Observation 30bb9c35-e5b1-4e07-b79f-2aa13f2e0a54 · outbound

This paper cites Sageattention2: Efficient attention with thorough outlier smoothing and per-thread int4 quantization, 2024.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Sageattention2: Efficient attention with thorough outlier smoothing and per-thread int4 quantization, 2024

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.665175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.665175Z digest=sha256:6cca8aea0568eb179ee404ea26571492aea987f365f07fda149c7db6bac175d5

Observation 5a1a2438-a6e0-474c-9e6d-6e00c78a1d08 · outbound

This paper cites Sageattention: Accurate 8-bit attention for plug-and-play inference acceleration.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Sageattention: Accurate 8-bit attention for plug-and-play inference acceleration

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.694887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.669290Z digest=sha256:aad8b941498c77f04c4faa710f9a5abd6b2006bed9b4ddeb5b0e3c2b8255ba3f

Observation 085d6e35-f4b5-497e-82e9-d20780ae7208 · outbound

This paper cites Spargeattn: Accurate sparse attention accelerating any model inference.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Spargeattn: Accurate sparse attention accelerating any model inference

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.673385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.673385Z digest=sha256:cd43bfa919424c1f0b1fb8c265b3c43fa73ef84e2ee55046a409095180a1b6f6

Observation a04e3274-0c66-45f4-933c-13d5e8ea6f21 · outbound

This paper cites A., Shechtman, E., and Wang, O.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity A., Shechtman, E., and Wang, O

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.677519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.677519Z digest=sha256:29d576ec9ce70e335140f81814c27381e9105d707ac89cd78301b75296ada680

Observation 3fe945a9-7ec6-4800-a241-5125d9efe6a1 · outbound

This paper cites H2o: Heavy-hitter oracle for efficient generative inference of large language models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity H2o: Heavy-hitter oracle for efficient generative inference of large language models

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.672309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.681700Z digest=sha256:55970f46125b0ab5ebabd0d449a47d4156579fa51b5b9a739f9dfbd27a6ee427

Observation 7b0c288f-7d89-415b-8072-19cf2c1705af · outbound

This paper cites H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.685957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.685957Z digest=sha256:cd85d73c2fb1cfed7c22e15b8af95061c6b90b523479f3089391ab68ebb20b54

Observation d54927b3-bdd1-42a5-85cc-fb8b08cd906d · outbound

This paper cites ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.690473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.690473Z digest=sha256:6c56915fffc324946528c4b88d48d6fe41a76526764c482c46ed3f9a5e4f300c

Observation ab7559ef-dab5-4f84-a086-e5305e5ff3ba · outbound

This paper cites Real-Time Video Generation with Pyramid Attention Broadcast.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Real-Time Video Generation with Pyramid Attention Broadcast

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.694822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.694822Z digest=sha256:85cf5f8be433c057ac7a160b881eb2f70df768915533af366c2e29d03020336f

Observation d96f1682-3934-4707-8d63-810069de4605 · outbound

This paper cites Atom: Low-bit quantization for efficient and accurate llm serving.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity Atom: Low-bit quantization for efficient and accurate llm serving

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:35:11.657809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.699192Z digest=sha256:171aaa5271389d9b7aa4c318e0b06bbe22f41e168a8782c464a054ac14580a26

Observation 4f3da14a-5d25-4838-8f36-a5e354cf739e · outbound

This paper cites PIT: Optimization of Dynamic Sparse Deep Learning Models via Permutation Invariant Transformation.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity PIT: Optimization of Dynamic Sparse Deep Learning Models via Permutation Invariant Transformation

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-08-09T14:35:10.761998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T14:35:10.703355Z digest=sha256:2f952c0e512ace4ad898da4370505554285154cd970a6b06eb42377175d4c1af

Observation 2b89fd62-6c36-4fd9-9b24-3cb4be2f45c9 · outbound

This paper cites write newline.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity write newline

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.707673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.707673Z digest=sha256:0bb059041b2e1853669bca8dc2e57e5ca5a7e974954cce0b4badd768ece275b9

Pith citing papers

Observation 11aba5e0-8fb6-4d8a-83ff-7e7460eebaaf · inbound

VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate cites this paper.

VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T12:38:39.235933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:38:39.235933Z digest=sha256:d05f861ae2a6b38ddd5409898ebd8ecb2e30f3dd4560f65ea30c42771320ba3e

Observation 49a85e5d-9b85-46d0-a13d-03e1082d3333 · inbound

MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention cites this paper.

MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T11:15:26.816861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:15:26.816861Z digest=sha256:aa2e903745db34bcc029441703ef677fa5d65ddc6f6c57d611ebfce741127b92

Observation 7f4b66ac-934e-4ac3-8294-58216d9ac79a · inbound

Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light cites this paper.

Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T11:04:02.010084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:04:02.010084Z digest=sha256:9b66123dcf6d98e6f288c90aad3147b82211e9a3e2b923ee41ed5c43d98e56e5

Observation 597935ef-4702-4e8c-b2f5-bb0c2900a916 · inbound

DraftAttention: Fast Video Diffusion via Low-Resolution Attention Guidance cites this paper.

DraftAttention: Fast Video Diffusion via Low-Resolution Attention Guidance Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:50:31.042863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:50:31.042863Z digest=sha256:030d1aa793ace14537c3002feca9b82243b70dd1777795f9690171965831623e

Observation 176d6473-8f57-4228-8406-b3946aca6f1a · inbound

FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge cites this paper.

FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:59.174155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:59.174155Z digest=sha256:cd7bd13ce8ad97e97063cf1e9e0c4cc564960025e3da85d4a8e8038abcdd2d40

Observation 28a836d7-a604-44bd-b446-d255f7676526 · inbound

RainFusion: Adaptive Video Generation Acceleration via Multi-Dimensional Visual Redundancy cites this paper.

RainFusion: Adaptive Video Generation Acceleration via Multi-Dimensional Visual Redundancy Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:11.296944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:45:11.296944Z digest=sha256:96c3bb298e54a9e4a5fbfb0edda6d0a6f0a43025ae47f7c3c4d5dd9717853e4d

Observation e310da7f-e507-47da-9b95-4f8ec4dc6b43 · inbound

SageAttention2++: A More Efficient Implementation of SageAttention2 cites this paper.

SageAttention2++: A More Efficient Implementation of SageAttention2 Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T13:44:50.914597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:44:50.914597Z digest=sha256:949e861f2a5d70b88d0b72ddb3658c448e8bdf274fedb409265f340297c0792b

Observation e7e096eb-6989-47c4-b97d-ae330f0af6cf · inbound

Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers cites this paper.

Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:17.692975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:17.692975Z digest=sha256:005d0e81f52fd8fad7f18c35d3191d71964c2c55cf6527e5f196ba2d2fa66a03

Observation 8888815c-1f7c-436d-bf57-7b8b239f0207 · inbound

Dual-Expert Consistency Model for Efficient and High-Quality Video Generation cites this paper.

Dual-Expert Consistency Model for Efficient and High-Quality Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:09.579398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:09.579398Z digest=sha256:3ae9e2e5d61258d35fc02952add62250983ded21d234320b03ff70089507cfcb

Observation 0f8e8662-d508-44ac-99ac-8055f127d9e2 · inbound

Chipmunk: Training-Free Acceleration of Diffusion Transformers with Dynamic Column-Sparse Deltas cites this paper.

Chipmunk: Training-Free Acceleration of Diffusion Transformers with Dynamic Column-Sparse Deltas Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T11:18:58.832753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:18:58.832753Z digest=sha256:52daace320f311aeaf22f1b8c29b8ed411dc7f0681fc00c929c4237b5a333f73

Observation eb3176f0-b26e-4050-811a-2033d5ee0502 · inbound

PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models cites this paper.

PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.488394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.488394Z digest=sha256:2178dc82760943919384f160eaa28a3ebd66407f4c807701f73900b966b50b4d

Observation 717e076e-05f6-4bc1-a421-d25a846e2616 · inbound

TurboVSR: Fantastic Video Upscalers and Where to Find Them cites this paper.

TurboVSR: Fantastic Video Upscalers and Where to Find Them Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:41:37.039836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:41:37.039836Z digest=sha256:9ab2acfd1bf99214c339def92443442449c312d7bebc091070960d68098a26b1

Observation bf59e288-1278-4493-969b-8c959b7228cc · inbound

VMoBA: Mixture-of-Block Attention for Video Diffusion Models cites this paper.

VMoBA: Mixture-of-Block Attention for Video Diffusion Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:05.295827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:05.295827Z digest=sha256:5106a56280ac73e4a139f470ba77f7bc194c2f39eb047ebf94202667a810b18f

Observation 44e63883-ca03-4026-8c47-1279fea49fa5 · inbound

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion cites this paper.

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T21:28:21.806986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:28:21.806986Z digest=sha256:1499c56003ad61b2d06349ae1ba54a33bb5bf52b67c82257671939761c084ca9

Observation 488f357b-a0cb-44cb-a952-f1d1a14d6909 · inbound

NABLA: Neighborhood Adaptive Block-Level Attention cites this paper.

NABLA: Neighborhood Adaptive Block-Level Attention Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:26:45.392305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:26:45.392305Z digest=sha256:d16a2e6847a17b723cd7e9b93212e77e02c1992dcc578b6dd6993b54be773043

Observation 7dca819a-450b-4f85-a5b3-2667b5231bc4 · inbound

TaoCache: Structure-Maintained Video Generation Acceleration cites this paper.

TaoCache: Structure-Maintained Video Generation Acceleration Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T17:33:44.566564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:33:44.566564Z digest=sha256:237f748b90f32cc780914ba38778f896615cecb5dd227e4f7bb280cf1967d7ac

Observation 12b00058-715c-45e1-8ef1-64774960aa2a · inbound

Waver: Wave Your Way to Lifelike Video Generation cites this paper.

Waver: Wave Your Way to Lifelike Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T17:45:18.378641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:45:18.378641Z digest=sha256:5b9c9b994c53da4345f8438b1cbd0c24b96fc291149b9ff3906d888527d400a3

Observation 2ffb7943-4b97-4d80-80f5-79b3fd6fc82c · inbound

OmniCache: A Trajectory-Oriented Global Perspective on Training-Free Cache Reuse for Diffusion Transformer Models cites this paper.

OmniCache: A Trajectory-Oriented Global Perspective on Training-Free Cache Reuse for Diffusion Transformer Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T17:42:59.370738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:42:59.370738Z digest=sha256:9bd51d3274c0b3a3cf8fcde4e99df8f78f83174ac36b72a3f06371b7daef6078

Observation a0c6e77e-7b67-44f3-8d44-0d7380e0be90 · inbound

FG-Attn: Leveraging Fine-Grained Sparse Attention in Video Diffusion Models cites this paper.

FG-Attn: Leveraging Fine-Grained Sparse Attention in Video Diffusion Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:02.071295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:02.071295Z digest=sha256:ed379715842f06e3053433ddce50bbe05706ca214f06ff838d941f656dd1dbe3

Observation 5f5d2aef-9f1a-4560-b2ab-e7ddb075913f · inbound

Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution cites this paper.

Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:51:20.342942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T11:47:22.725217Z digest=sha256:9f04fd68fb864ebdaf129dd61b3a0ca1d963d059725036c9a5944f8cca283721

Observation fc7f59fe-43c7-43d8-aebd-faa67ca10143 · inbound

UltraImageGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention cites this paper.

UltraImageGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T09:20:07.411349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:20:07.411349Z digest=sha256:23ebc8c1d6a6e14a16fd22175707f11445e11e5f2c51a87ba7e337fe6f3a44d3

Observation 29858386-6753-4d0a-b7d9-e4e9a1103b10 · inbound

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space cites this paper.

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T22:12:45.435647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:12:45.435647Z digest=sha256:bc38201f8d5a0b3f2f39a86b2686c3b8732f67272fa85a666210eb1315c7362c

Observation 7d0d5553-bf7c-4ced-bdb5-657b68c2f09f · inbound

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling cites this paper.

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-03T15:46:12.720965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:46:12.720965Z digest=sha256:f22bc9a6bfd219a4e5c05be34ef06e94652162c18cb103d7369e0797bb31b21f

Observation c3ebcfb7-2543-4843-82e6-59217e08f0c9 · inbound

Trainable Log-linear Sparse Attention for Efficient Diffusion Transformers cites this paper.

Trainable Log-linear Sparse Attention for Efficient Diffusion Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T15:35:16.759708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:35:16.759708Z digest=sha256:04abee5de6e79417ea9cc6da0d3923215496be7982326e365defd9e6f064740b

Observation 40e70530-067d-413c-9cb8-5e8c17ed031b · inbound

TinyHistory: Lightweight Video History Embeddings via Two-Stage Context Learning cites this paper.

TinyHistory: Lightweight Video History Embeddings via Two-Stage Context Learning Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T13:35:50.008026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:35:50.008026Z digest=sha256:924bb657e5cb11754d3a23dac4a5e2da0b5d7789cc718616a6a51823e7a32f4a

Observation caa6157e-9703-4170-913b-e912246b9dfc · inbound

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices cites this paper.

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T10:56:16.020021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:56:16.020021Z digest=sha256:e89272065bf0cddcc3f1b189ba16666339a91c1dfe5a788d4d5dbe723e08e8bb

Observation 03b351ad-57de-4fe1-ae9d-5270f8adcd29 · inbound

Transition Matching Distillation for Fast Video Generation cites this paper.

Transition Matching Distillation for Fast Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T10:35:06.618716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:35:06.618716Z digest=sha256:12cbe848555fb7921071288adac23d9cd616663b3132306c0e78265fea05b257

Observation 12119fda-2811-4a23-a6ff-7627ad1cdcea · inbound

Mixture of Distributions Matters: Dynamic Sparse Attention for Efficient Video Diffusion Transformers cites this paper.

Mixture of Distributions Matters: Dynamic Sparse Attention for Efficient Video Diffusion Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T10:37:47.346990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:37:47.346990Z digest=sha256:b624b0f9594f893d5fd0ecb25cb34b706685ada047530e8a50ccef00e17864ef

Observation 79c64972-fb6b-44b8-9633-1f12f4bcbe5d · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:32:01.847359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T17:32:01.642256Z digest=sha256:0724ba14ba6f298dfd99cc5eeb3474fc1e836a4181c89edf495436936ad64c4a

Observation b50e13c0-b11d-4195-a62f-4b639e46a6ed · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:51:29.979301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T11:48:00.633421Z digest=sha256:90c92731e678e4e92f48a65e866571543deb907f89f516840ceab1dd6d0f2524

Observation fbebeb35-b960-4f6f-bae0-f6c268a5e08d · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T05:30:28.604102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:30:28.604102Z digest=sha256:e9124611a9404dab2c12ae39035f13ca6eb8d7614977c4ec76ba6287fe28e053

Observation b14e174c-2216-42b4-9311-00c1c05a1efa · inbound

Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization cites this paper.

Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:00:44.646845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T07:58:24.456859Z digest=sha256:c19c6773a828f356033d7bd73d97b3606ea472c6e9a8d6687616a7af1c8e6cbe

Observation bd3b27e8-69d2-4be7-8ddd-69213ce2fd4c · inbound

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention cites this paper.

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T04:33:11.377277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:33:11.377277Z digest=sha256:ae5b23ca8c0a7fb6d9bc0afc43690fcb1d00ccbfe6e3f5dfd024e3911ebb3c35

Observation 2e0fb8c7-50a5-40f6-a6dd-35b2376f3d78 · inbound

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention cites this paper.

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T00:38:17.262351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:38:17.262351Z digest=sha256:604c4ea01cd6c199dd106927f503948577c6d998352934706472a9ab9f9340a9

Observation ecaa71af-89a8-40c4-a854-a7c060b3b420 · inbound

S2O: Early Stopping for Sparse Attention via Online Permutation cites this paper.

S2O: Early Stopping for Sparse Attention via Online Permutation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:36:32.881938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T19:32:52.948154Z digest=sha256:788a5879a61e5be5508d1569d994082f445ef212017a4724a34936923cac6e3d

Observation 237d2c6b-f340-43d5-94ce-b73109df66cc · inbound

Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery cites this paper.

Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:06:14.412858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T16:05:36.307084Z digest=sha256:81fe7e94639e7d1b80647058315c18258d04c7d8b6ea190fe2b7afc7389526cf

Observation 033b80b0-a7a8-4f85-8da4-f7a1b3b8e1b0 · inbound

Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering cites this paper.

Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:59:53.189960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T08:55:52.757489Z digest=sha256:f5de52096a5ea89f8e86e53468d4ac5d431486cec3fa00ea99c7bb18c2ad2528

Observation 0e6c0224-9558-48ca-bf5d-0bd1bc36ccfb · inbound

SURF: Signature-Retained Fast Video Generation cites this paper.

SURF: Signature-Retained Fast Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:14:17.576394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T18:11:39.642701Z digest=sha256:d73f27fcedbc660dca1b2dc09f2e9f8638c04ba46c35d553d12dce44ca9611cb

Observation fd7d2427-4387-4476-b61f-0a66bdd30ead · inbound

Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation cites this paper.

Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:11:10.460354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T16:10:54.363560Z digest=sha256:68b4ca6d194519b2644ad1e51d2b99ae321331a3136130d77ae9fb4e452e066b

Observation f8708ad2-5f18-43fa-ba59-6d9c95b2a0ec · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 152

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:25.714069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:30d7cef2dfa418c17ac63c3a66fec53020499f87717c39c13b3ef18f23522a32

Observation f685719a-ab53-4a16-b658-985b3581a6d3 · inbound

Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation cites this paper.

Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:35:19.138279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T04:49:44.525757Z digest=sha256:3efa7284cf5ccf4c7422225fe26fa815d85d3794ded4e241d87f2943539e6dfa

Observation 15acc78b-a1d9-4421-9b00-e34d473a2ca2 · inbound

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation cites this paper.

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:49.671215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T05:09:46.155328Z digest=sha256:1b90a2743e3c237bfb61501a6cf36bf96ca1a6102488d0c02a84c775c4c1dce8

Observation 86555e28-e483-46d4-9c50-672909cb7cf1 · inbound

DynamicRad: Content-Adaptive Sparse Attention for Long Video Diffusion cites this paper.

DynamicRad: Content-Adaptive Sparse Attention for Long Video Diffusion Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:07.908341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-10T01:25:27.819753Z digest=sha256:6c2fe393d79ea9c39daef0d5c4ed237e13eb062e0ec6f4619c1c673edd038efc

Observation 65b90f58-c4ac-4e88-bdd1-7390145cb356 · inbound

CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers cites this paper.

CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:49:44.342655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T04:47:32.476614Z digest=sha256:983df35b3fc89809ff8592e49723dbd22a0497e554d780a449a89d993cfbf1cd

Observation 5f9cecb8-8c3e-4cb4-9cc4-a890597178d1 · inbound

HEART: Exploiting Head Heterogeneity in Sparse Attention for Video Diffusion cites this paper.

HEART: Exploiting Head Heterogeneity in Sparse Attention for Video Diffusion Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:19:46.250983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T05:17:48.403406Z digest=sha256:7a7a6bf211b4dd431d373669cc76803ced51be02228348f40dfc9224b49fcd6e

Observation ed1dc2ec-a890-4357-8866-461fcf7ce513 · inbound

SparseSAM: Structured Sparsification of Activations in Segment Anything Models cites this paper.

SparseSAM: Structured Sparsification of Activations in Segment Anything Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:48:19.746267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T13:45:26.106499Z digest=sha256:15a918a70d82ee330a0aa4538c797c453ab09e1606c5019bdc95789772ca0ab2

Observation 64e07104-0506-4a3f-8e73-5dabbc86bf40 · inbound

DFSAttn: Dynamic Fine-grained Sparse Attention for Efficient Video Generation cites this paper.

DFSAttn: Dynamic Fine-grained Sparse Attention for Efficient Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:45:20.097693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T04:44:19.628926Z digest=sha256:ba721c2725705575b5059acd097e37a53c44763e24cd526039b10888854b8e19

Observation 33bf7c9d-bf25-473a-bfa4-4579457705b9 · inbound

DiSC: Resolution-Scalable Acceleration of Diffusion Models by Exploiting Sparsity and Cached Token Reuse with Hash-based Distribution cites this paper.

DiSC: Resolution-Scalable Acceleration of Diffusion Models by Exploiting Sparsity and Cached Token Reuse with Hash-based Distribution Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:33:53.916733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T19:33:05.248105Z digest=sha256:d732d927dd2f3c19231334c37c790aa36e01c30756bc5f1d609abf2990eead46

Observation 282da279-6c72-4d83-8773-c6ced65095f5 · inbound

RT-Lynx: Putting GEMM Sparsity in the Right Place for Diffusion Models cites this paper.

RT-Lynx: Putting GEMM Sparsity in the Right Place for Diffusion Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:43:54.747044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T19:40:42.033793Z digest=sha256:36ad80ffa42733af257815a40683616fb03827255f4a3cdbb6eb32e37e52cd64

Observation d4b1cc71-f3b3-44e6-b793-4447d2ac072f · inbound

OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning cites this paper.

OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:13:27.533279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T13:06:02.201607Z digest=sha256:07dbba5b0a1e45ac74129066e69372846ecfd0d5613be2286471218b2dbcce1c

Observation 99cdccd7-71f1-408e-a760-0cf1e75ddf3f · inbound

Veda: Scalable Video Diffusion via Distilled Sparse Attention cites this paper.

Veda: Scalable Video Diffusion via Distilled Sparse Attention Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.720126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T07:54:13.390911Z digest=sha256:0af7df82dfa1646e38570ed384f24add888f9515e42c1bb1092ddbdd95bb38e4

Observation f6732574-dd6f-4a64-91b2-1727272b69f0 · inbound

LVSA: Training-Free Sparse Attention for Long Video Diffusion cites this paper.

LVSA: Training-Free Sparse Attention for Long Video Diffusion Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:45.076275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:35:11.610183Z digest=sha256:779c02130f9dab10d1594ecd594776fdb245724e363e8424cc633844933ba5a9

Observation 79586b40-ad51-4cad-9c63-c8ea058d3fc4 · inbound

Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models cites this paper.

Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.145516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:43:32.749425Z digest=sha256:482564cca722b2240f75b8542deb758b0f04de82371f7e17694c011b3356b3db

Observation 7e67f4da-b947-4130-a9e0-0ed777b9f8bf · inbound

RhymeFlow: Training-Free Acceleration for Video Generation with Asynchronous Denoising Flow Scheduling cites this paper.

RhymeFlow: Training-Free Acceleration for Video Generation with Asynchronous Denoising Flow Scheduling Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:56.724770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T02:08:44.257127Z digest=sha256:5cec41983a9774d51f12c89089a6b7fd317314946484733d8cf3e3008fb536e3

Observation 800a0f5a-6871-4465-85f9-71a6ab88187a · inbound

ScalingAttention: Discovering Intrinsic Sparse Attention Topology for Video Diffusion Transformers cites this paper.

ScalingAttention: Discovering Intrinsic Sparse Attention Topology for Video Diffusion Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:49:44.816794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T09:22:55.731241Z digest=sha256:251b21898c95112e013026b1a9fb073218bc08256ef66adb1aaf4cc78e8ec032

Observation f30539c6-8789-4c39-9d06-b6a6890d702e · inbound

Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation cites this paper.

Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:49:41.596153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T11:05:44.972128Z digest=sha256:e2ff6f069011ff7d94b45af42155f3c09227aa6b0ccf4ed25018dc1edb0822d5

Observation 9e76f868-216d-4563-a5d4-88eeb61ebbbe · inbound

EcoVideo: Entropy-Orchestrated Video Generation Paradigm in Cloud-Edge Dynamics cites this paper.

EcoVideo: Entropy-Orchestrated Video Generation Paradigm in Cloud-Edge Dynamics Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:14:22.140642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T06:15:31.587641Z digest=sha256:16053006564581069efaa2ee6ed14e84c6357b58bfd546b2ac0907b8479d53b9

Observation 34383717-cb27-4789-97c3-20285c38ae0c · inbound

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation cites this paper.

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:45:40.316344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T06:13:11.504526Z digest=sha256:e88156d2bd3299241f40fd0d9ac76202870e5849ff8f44acb862fdb0c129690b

Observation 7acff67a-d000-43a3-9681-d355d9067ea5 · inbound

SAF3R: Dynamic Sparse Attention for Feed-Forward 3D Reconstruction Transformers cites this paper.

SAF3R: Dynamic Sparse Attention for Feed-Forward 3D Reconstruction Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T01:12:13.747675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:12:13.747675Z digest=sha256:1d69ad75e6b315df2e4dd426aa75b4770d99c4b101f6dfb0500b4b37890a981f

Observation 3480828d-6d6b-4a0e-a372-973e60cbff29 · inbound

Controlling Motion Transfer in Diffusion Transformers via Attention Heads cites this paper.

Controlling Motion Transfer in Diffusion Transformers via Attention Heads Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-14T07:11:45.493916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:11:45.493916Z digest=sha256:59c3d0e623646fa42662244ed3f293a0a9ebb0fc5fd6e44155897820704a2ef2

Observation 25f0dc1f-cc73-47b9-9044-a0c35f4a2caa · inbound

ACID: Adaptive Caching for vIDeo generation cites this paper.

ACID: Adaptive Caching for vIDeo generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T06:40:16.756592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:40:16.756592Z digest=sha256:46fb8f251095b2b3245400894f431d26723c6e5d75eedd4ac955fd54c6a02033

Observation 85664d8c-dcf4-4452-92cc-75a0efeaaac4 · inbound

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations cites this paper.

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-02T03:51:52.201118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:51:52.201118Z digest=sha256:322160f253a7b3c54f7e7c71eb0761b48644d7ea89e443f1565f5390e0c7175e

Observation 96c6444f-561d-4539-8dc0-c48ba40df209 · inbound

DiTango: Cost-Effective Parallel Diffusion Generation with Selective Attention State Reuse cites this paper.

DiTango: Cost-Effective Parallel Diffusion Generation with Selective Attention State Reuse Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T22:45:28.423325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:45:28.423325Z digest=sha256:c05f1e1c877fb91cfed6b0258d5a961a3bc46a14dfe2e83ad94c294a47261e8b

Observation 71237a44-5ff7-43b0-90c8-cd58819ff435 · inbound

Surprise Forcing: What to Remember, When to Skip in Long Video Generation cites this paper.

Surprise Forcing: What to Remember, When to Skip in Long Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T15:29:45.954883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:29:45.954883Z digest=sha256:7482a1f100fbdef056cb735fb53c29dbb4731f78b7ceae3ff3b0834f6453efa0

Observation 9653e097-2e9c-4e36-a90a-2cbb82949270 · inbound

SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation cites this paper.

SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T07:09:08.982968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:09:08.982968Z digest=sha256:8eda9da7b6a6a0bb9461976327d69f6954af181d8b96916e0f6665fcfdf18dd9

Observation d4b9599c-978b-4803-9284-d698d4c5d12c · inbound

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models cites this paper.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.811138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.811138Z digest=sha256:a698e46ad1f076bb83607ad33f78faf006f569664283e0dac58cc0966fe098b1

Observation 9f33261d-732d-49fe-878b-39018db1bfa9 · inbound

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers cites this paper.

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 89

Resolution
unresolved
no resolver link, observed 2026-07-31T02:16:08.242540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:16:08.242540Z digest=sha256:ef838306b0b42fac08f109fdf8ab45b88cbf85d56fe25c0eddf77dee10dc237f

Observation e0991f63-ee9f-4291-9d57-b0a7d8a4bdf9 · inbound

EchoCache: Energy-Guided Cross-Modal Caching for Efficient Audio-Driven Video Generation cites this paper.

EchoCache: Energy-Guided Cross-Modal Caching for Efficient Audio-Driven Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T06:45:49.417885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:45:49.417885Z digest=sha256:939ed513e1a50a31fb3cfb47c6b6906417218428b6d678212037fdb924c68b84

Observation 50cf944d-afd2-459c-8591-a4feb3e27e20 · inbound

Token Radius Attention for Efficient Video Generation cites this paper.

Token Radius Attention for Efficient Video Generation Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-04T06:03:07.185566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T06:03:07.185566Z digest=sha256:5f3f55c306b7d99b390c0070abdacc3b55d53685bcf44228f4224189f5d00fe8

Observation 2fb1c392-37b1-432d-936b-09a318b8a644 · inbound

SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inference cites this paper.

SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inference Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:50:04.013151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:50:04.013151Z digest=sha256:6c876558e4dd3ec60275f3d4fa8fbe2758ca5ddc947898b5075ce816afaa2b06

Observation e7e5fb4a-b431-43a4-b2d3-478cf5288061 · inbound

Deltoris: Enabling Real-time VLA Inference in Embodied AI via Bit-level Sparsity and Speculative Inference cites this paper.

Deltoris: Enabling Real-time VLA Inference in Embodied AI via Bit-level Sparsity and Speculative Inference Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-15T14:46:04.288262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:46:04.288262Z digest=sha256:93d37309647c1a56b4fe41fdabf7784363062faf83f87b05b7c6e203b798d9a4

Observation 6f7f4991-0c42-4061-b600-9f7b66cdf7cf · inbound

ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing cites this paper.

ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T13:04:14.205005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T13:04:14.205005Z digest=sha256:e77730e0c8126dcaa46b9a10fabb1a5c45f119bda4ca32e51eeb8e722333e5d2