Pith. sign in

Paper Citation Record · LEDGER

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

As of 21 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 33 inbound Pith citation observations for arXiv:2411.16375.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16375 v2

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:17:41.946523Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 33 of 33 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:40:46.929553Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T14:03:48.589989Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 194c750c-ccc7-4468-a694-67d33009b31f · outbound

This paper cites FlowZero: Zero-Shot Text-to-Video Synthesis with LLM-Driven Dynamic Scene Syntax.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing FlowZero: Zero-Shot Text-to-Video Synthesis with LLM-Driven Dynamic Scene Syntax

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.872257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.872257Z digest=sha256:5228b6c0673b753d22e2ee47616b026896905d8842251eac1d7ff4b9744904e2

Observation 6234e286-06c3-48d7-a0c4-cb071354b5e5 · outbound

This paper cites Videofusion: Decomposed diffusion models for high-quality video generation.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Videofusion: Decomposed diffusion models for high-quality video generation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.250515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.877000Z digest=sha256:77751f2a8412b8dd8f4dd017c323b7c9ec1314f5ae63a939eb1073ed1ae6bc9d

Observation 4ba91052-d0b2-4530-b6c4-1e5368d5aeab · outbound

This paper cites Stylegan- v: A continuous video generator with the price, image quality and perks of stylegan2.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Stylegan- v: A continuous video generator with the price, image quality and perks of stylegan2

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.238276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.881586Z digest=sha256:b95134d311b8c1190d84247f8519d71178bd6868fc0183dd6710f6194da94940

Observation 6487a23f-d84f-4c4d-9561-9a022dbf9b4e · outbound

This paper cites Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.895638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.895638Z digest=sha256:f4d738772584f57eb4bb7b86320691c713bc8f2de033665620b0cf8dab9fa9fe

Observation c641ef26-45f0-4153-b476-ae169edd4f44 · outbound

This paper cites Live2Diff: Live Stream Translation via Uni-directional Attention in Video Diffusion Models.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Live2Diff: Live Stream Translation via Uni-directional Attention in Video Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.900937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.900937Z digest=sha256:c36106cd34e71084a660d327c506253f9ed18aec34df33516e39d59a7da0b830

Observation f71f2c5c-8d78-4e8a-9b06-5f24c7ced382 · outbound

This paper cites Detailed Training Objectives Recall that (cf.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Detailed Training Objectives Recall that (cf

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.213116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.912253Z digest=sha256:28f517da8f08dee8f4c39ee12962ea7f1a2722df575cc7e19fade3ec9a6751cd

Observation 4b9ec7c7-b9a2-4ced-8f22-ec60238c960d · outbound

This paper cites (6) Since q and pθ are both Gaussian, DKL is determined by the mean µθ and covariance Σθ.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing (6) Since q and pθ are both Gaussian, DKL is determined by the mean µθ and covariance Σθ

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.200788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.916414Z digest=sha256:d62ad369800bc85cc8f9473b524e4ec6051a93a29d5d61ebf6f10224d0ca0882

Observation 140c8a1b-b74b-4553-9e53-33a4d52b8239 · outbound

This paper cites UCF101 (Soomro et al., 2012).

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing UCF101 (Soomro et al., 2012)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.174766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.926030Z digest=sha256:73bfc78b497cdbdf7ddd6a78987b98e5d65ade01e99cead8ae5d3a46aadd868a

Observation d94123b4-342e-4fbb-91a6-6ef3eba00802 · outbound

This paper cites the first AR step.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing the first AR step

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.149972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.934408Z digest=sha256:8fedda0dc7fefa44e642b5a14854c44169ae072c875e76f2ecb9997f99c61560

Observation 691f1efa-f1b2-4484-854d-4a19ac3262fb · outbound

This paper cites VBench is pri- marily designed for text-to-video evaluation.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing VBench is pri- marily designed for text-to-video evaluation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.136769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.938298Z digest=sha256:816e87e914bc27d6dbddd79704a395d08c95486bb34ea9853d465c37b4c7ac3f

Observation 6272dc5d-cf28-43a0-bb03-4a95bc4181a2 · outbound

This paper cites an unresolved cited work.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:17:42.104629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.946523Z digest=sha256:aea9497d30c54047b9444776a2653b7b885e88a9d397f27b66baf3f144e80de5

Observation d5fe8987-d9ed-498f-b2b1-eafd2b7f8cba · outbound

This paper cites Method Aesthetic Imaging Motion Temporal Quality Quality Smoothness Flickering OS-Ext 44.39 50.74 98.93 98.57 Ca2-VDM 44.30 50.55 97.59 97.14 OS-Ext baseline.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Method Aesthetic Imaging Motion Temporal Quality Quality Smoothness Flickering OS-Ext 44.39 50.74 98.93 98.57 Ca2-VDM 44.30 50.55 97.59 97.14 OS-Ext baseline

Reference 256

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.123203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.941929Z digest=sha256:95a4d99f387c047a755a5cbb2acf0837a4807b7b2e4f8af04452fffc5fd18af5

Observation 25028251-2f52-4166-a5f7-5deb884ed57d · outbound

This paper cites Fvd: A new metric for video generation.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Fvd: A new metric for video generation

Reference 2012

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.226059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.890848Z digest=sha256:e805c7ae8851fd7de9556c93c13b76053cd982049c052949ade63ed22c8b2146

Observation 00316872-6dee-4c52-8ac8-59f10b20f853 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.886233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.886233Z digest=sha256:c5b6056cb05bae3ef51c81db11e74ebd288e223c091991d78e8329bfac5f61d4

Observation 6952da28-5f11-486c-a47a-a2b98f5c3a03 · outbound

This paper cites We followed prior works (Blattmann et al., 2023b; Ge et al., 2022; Ren et al.,.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing We followed prior works (Blattmann et al., 2023b; Ge et al., 2022; Ren et al.,

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.162530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.930503Z digest=sha256:8873f2f58dc0134d7fa5694d0661bec540a1340d4e84578643ca1e822574e243

Observation c4c143c6-06d0-4e4f-8f81-814009148a4d · outbound

This paper cites I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.906371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.906371Z digest=sha256:e6706ac3b2f461b07cca6fd2df52c94d91c6525459abb8f5d712bcc5976020fe

Observation 3e2575e0-735c-4a0b-954a-33c95ab3eaf0 · outbound

This paper cites A white tiger in a zoo swimming in a lake….

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing A white tiger in a zoo swimming in a lake…

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:17:42.187964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T13:17:41.921214Z digest=sha256:797829707507ca573bfa1f6455ac6fc05a32a9b42bd8f21cd5517d079571436f

Observation 78c53871-ec54-46a2-8ef7-7592e89ca92c · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.861604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.861604Z digest=sha256:df61b3c59bd51f856e23b0780d73f2a37ec1b1b409d0b1568c9a0245ac596756

Observation 968da582-aa1a-404c-bb2c-3491e5f0e446 · outbound

This paper cites StreamDiffusion: A Pipeline-level Solution for Real-time Interactive Generation.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing StreamDiffusion: A Pipeline-level Solution for Real-time Interactive Generation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.866200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.866200Z digest=sha256:06d35dd2c3153dc279225a5ea988dee2b9016ebbe9170170d699ec30ce77391f

Observation 9f6ae92b-648b-417e-bf98-e2782f8b95d1 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.855795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.855795Z digest=sha256:fa7e9a71fc9fec2f7375f521a6debfdbc60b12003ad664f450f0bd2301be689c

Pith citing papers

Observation f233286d-4875-406a-85c5-81bd97b25165 · inbound

Long-Context State-Space Video World Models cites this paper.

Long-Context State-Space Video World Models Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:03:15.086030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:03:15.086030Z digest=sha256:611887ebd4398b1f3acb6932212ff34aae644d59a02df880125dead4402097ca

Observation 80aac502-dcbf-4189-9864-0b149aa2a270 · inbound

Video World Models with Long-term Spatial Memory cites this paper.

Video World Models with Long-term Spatial Memory Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:37.670335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:37.670335Z digest=sha256:7f6de29a4fb33448ea6643c7cf3e07298716b4b39d5db16a614769c4d724cb43

Observation 4f9735f1-ae02-45dc-8a3d-0ad4c05d0e04 · inbound

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion cites this paper.

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:36:53.142419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T01:36:53.029590Z digest=sha256:92784573dbae965ef22fb3ec563fe015a737ac3156ff8bd744c7b29c08409561

Observation bc9f614f-4af4-4006-b277-df5a71160220 · inbound

CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion cites this paper.

CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.775719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.775719Z digest=sha256:13711f5bc64cff710b16b44c896c0507108f5d32e50c09fdffaa593d2d0b7bf6

Observation d2d60708-f89c-49fe-9fa1-029466172e46 · inbound

LoViC: Efficient Long Video Generation with Context Compression cites this paper.

LoViC: Efficient Long Video Generation with Context Compression Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:23.020880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:39:23.020880Z digest=sha256:d087824f56433c1d3d9eec389187bced9b55c5be3b79d423845123ff6803eab3

Observation 9bc497ef-ca4b-4856-b249-1f3b39a4cda1 · inbound

Flow marching for a generative PDE foundation model cites this paper.

Flow marching for a generative PDE foundation model Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:51:25.606866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T13:48:14.532529Z digest=sha256:7d4ea1945b75705ef635ebdcfe2d4b65c29b00c904eef4aea67001936ceb3d2a

Observation 5b202109-498a-41b2-baf2-42ba1e416eb5 · inbound

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling cites this paper.

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T15:46:07.054788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:46:07.054788Z digest=sha256:864754f3f143b47d24863978205a3f11c84564c2421848153a6b67a596b21899

Observation e77cfe51-8f5a-4821-9221-44e927289657 · inbound

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention cites this paper.

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T04:33:10.693871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:33:10.693871Z digest=sha256:94bb4ffbae634a4f62f0c6c4b7ea22e81d067c08471a4dcb63b9c4c23b5c61b0

Observation c8362b4e-f1dc-4a15-b6e5-36187eccc7b1 · inbound

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention cites this paper.

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T00:38:16.421515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:38:16.421515Z digest=sha256:06106b21dd3a0ebc0e59585f39416d608c4814d60e4ca41b81788cf4eef20fa6

Observation ca5e67c7-1c22-4fc3-bb25-1b835f594f37 · inbound

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion cites this paper.

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:07:29.899500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T07:02:38.876518Z digest=sha256:1cad0afa70a8cb8b119e7e83e06fed272d58b3693c9f46b7d578cf273ed9888a

Observation 759246b3-d963-473b-a31a-e2b2d711c9ff · inbound

Latent Generative Solvers for Generalizable Long-Term Physics Simulation cites this paper.

Latent Generative Solvers for Generalizable Long-Term Physics Simulation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:22:22.564867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T05:21:13.280286Z digest=sha256:c2b19605cd8b086d0005933867b6b1b62684b4ddffa7e2fc5eec326998fe38c3

Observation 7fcb2080-e596-4527-b256-dea7d962cfb3 · inbound

World Action Models are Zero-shot Policies cites this paper.

World Action Models are Zero-shot Policies Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:18:15.925581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T16:18:15.003371Z digest=sha256:6a4dacc8cb77100f985b69706aa329d8f9749c01112946c222e172d3a4cdc00e

Observation d6b0969c-bd83-4a35-846b-a0c85c5cf219 · inbound

ChopGrad: Pixel-Wise Losses for Latent Video Diffusion via Truncated Backpropagation cites this paper.

ChopGrad: Pixel-Wise Losses for Latent Video Diffusion via Truncated Backpropagation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T09:45:23.285493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T09:42:26.074609Z digest=sha256:407c6442ca90c73851d4bd1b24f1637915d2a5efdf068ead28c65d4c6e8baeab

Observation 82e72f8e-d72b-4b24-8c4f-2e8523944d3d · inbound

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling cites this paper.

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:25:53.514242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T19:08:56.588282Z digest=sha256:0c722208ba6ca89b5d185d3af914a7f1d983501fb34be202e3f4309451bd2eaf

Observation 076b2ddf-a4f8-417b-ba6d-754b106dced9 · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 274

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:26.211296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:4253e10f44872f3692085b7eea19e1e4aa3bd2be9f9f1c7f5ee4f8e66f6ae19a

Observation 62b5c462-02c7-4d39-970d-8885c0eaeed1 · inbound

Stream-T1: Test-Time Scaling for Streaming Video Generation cites this paper.

Stream-T1: Test-Time Scaling for Streaming Video Generation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:40:39.381562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T18:16:14.985693Z digest=sha256:e40967fba54472c08d06c79f87a8f251929f13a45032a4eca9ece359544a46c3

Observation 0463eeab-4d57-4e39-b9f3-e3db5f6d7b28 · inbound

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity cites this paper.

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:33:32.281615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T02:31:48.593354Z digest=sha256:330cd8a903d69d4f630ad1018b36d0822112bfc66ffd22a41043750e3a2fa909

Observation 3301eccd-ccc9-44a7-835b-c2661441dc68 · inbound

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization cites this paper.

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.560960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T20:11:33.277349Z digest=sha256:79deb70c6ff0fc32fe614c6d8c8ae2bdddfe949ce47a9672544b406b80302d81

Observation 4c354c30-61e1-4c92-999c-6620fc5c2628 · inbound

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization cites this paper.

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:25:00.605936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T19:23:07.044659Z digest=sha256:1c20540e680a9798dc3151475a0c8547b27ee6afb2add7d351410f1b7e751088

Observation e4ea9939-6545-4d78-be35-9ee771891979 · inbound

HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videos cites this paper.

HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videos Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T13:33:18.992886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-20T13:32:36.375456Z digest=sha256:98cc14b8951909162123d839b6712634d5bdbe45c93a06efece0e2f751fd6b34

Observation a36d3999-c64e-4578-9e3b-d52981915024 · inbound

MORPHOS: Autoregressive 4D Generation with Temporal Structured Latents cites this paper.

MORPHOS: Autoregressive 4D Generation with Temporal Structured Latents Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:19.522972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T15:03:22.366584Z digest=sha256:d1564ad74da5f94957add41c5dadf107a5152620f1369bc7762510372ecfc135

Observation fa469b43-2e86-4cab-ad9d-db4477d7eb1c · inbound

Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control cites this paper.

Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:43:51.098440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-29T05:03:56.624536Z digest=sha256:18597c8b58940192d825ecdf00a15564e4fd3ad7d1c1cbf65cf67bcdaada64d7

Observation 57f5c1b1-aa72-436b-b682-d0ccee7287f9 · inbound

TempAct: Advancing Temporal Plausibility in Autoregressive Video Generation via Planner-Executor RL cites this paper.

TempAct: Advancing Temporal Plausibility in Autoregressive Video Generation via Planner-Executor RL Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.604394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T04:46:48.414652Z digest=sha256:c84e96dd6a1cd7107056a0b166483f01e62766686c22e546947a862134ec233b

Observation 4794075d-da37-469e-b6d8-acb19dd855b2 · inbound

TempAct: Advancing Temporal Plausibility in Autoregressive Video Generation via Planner-Executor RL cites this paper.

TempAct: Advancing Temporal Plausibility in Autoregressive Video Generation via Planner-Executor RL Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:39:00.994390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-03T22:38:33.557109Z digest=sha256:142d8893749b2d93fa0b90098e2de7c3e551ef04bd418b942ea2f686b58d23bc

Observation fcfb7cc4-8a79-4f78-b3e0-1d7df73b92b9 · inbound

DreamForge-World 0.1 Preview: A Low-Compute Real-Time Controllable World Model cites this paper.

DreamForge-World 0.1 Preview: A Low-Compute Real-Time Controllable World Model Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:34:21.442236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T07:32:03.567262Z digest=sha256:989453c80034c4aee903726211f667726cc0fe7f6e38f02331622e33c3338029

Observation e150c102-7f76-4788-b173-fb78bf046fbf · inbound

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation cites this paper.

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:26:53.857256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-02T11:21:32.725966Z digest=sha256:ac702277d24790cb0d2a0f138d820832c933b292e3ce68fea54afd76b72a1c32

Observation 37279416-c207-4d80-8e26-1e27e76ef556 · inbound

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation cites this paper.

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-14T16:48:30.512494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:48:30.512494Z digest=sha256:343989a9d72b5af1ba95b7dbdb632690348a65d69a7added57f545398d5435b5

Observation 1fb307db-2d00-4d64-90d0-1e596cf49863 · inbound

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation cites this paper.

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T15:40:46.929553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:40:46.929553Z digest=sha256:44f9118dd21f297ab78c353f3abcd6424e949b23552b8f59b36e6dc8c19c7b14

Observation c9783dbd-9e12-48b7-94f8-1d58a3c5ce84 · inbound

MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing cites this paper.

MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T14:03:48.593218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-07T13:58:13.194386Z digest=sha256:a2e728f6f1674a0f66b4b0ce1dfe82d22d6cdeeb755774b38cc47c8caec56698

Observation c3ac8b02-b32a-4811-80a2-7b0eab42580b · inbound

Wonder: Video World Model Done Better cites this paper.

Wonder: Video World Model Done Better Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T00:52:24.924518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:52:24.924518Z digest=sha256:8fa601f8c325ad7c967e6b7edab03abfc1647f1e8a58b5825d0fce9d0d12096f

Observation b5138675-3957-4e47-ba3f-3a01b3ffe7a2 · inbound

MiniWorld: Democratizing the Training of Video World Models from Scratch cites this paper.

MiniWorld: Democratizing the Training of Video World Models from Scratch Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:24.245292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:24.245292Z digest=sha256:9c74b7549b5b912895e2b177b4d1090ca35b958f52aebf88bea5ffe549315c46

Observation 9826d9b3-6cff-4d49-b97a-d4a3a2b57671 · inbound

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation cites this paper.

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:26:02.465168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:26:02.465168Z digest=sha256:f6c0c4e2f388b42e7d55b944ef91c2739d2a54dbc8a4242fa5061aa36d1b5950

Observation 6bfa884f-388a-4dc2-948e-d44ba3e69bdc · inbound

Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation cites this paper.

Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T12:17:52.836008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:17:52.836008Z digest=sha256:987b29cb2f81565e4eea95e963102e11872dc75898e7c293c78b3c856e308ec7