Pith. sign in

Paper Citation Record · LEDGER

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

As of 6 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 100 inbound Pith citation observations for arXiv:2310.19512.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.19512 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-14T21:40:43.956642Z

measured 163 of 163 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 100 of 101 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:39:21.916399Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact24
  • verified fuzzy34
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

37
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation d842b9f8-4ca1-438a-96b6-f4a5af2a35ab · outbound

This paper cites Accessed October 22, 2023 [Online] https:// research.runwayml.com/gen2.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Accessed October 22, 2023 [Online] https:// research.runwayml.com/gen2

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.263992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:a3ea01640d850e3e9ceff40f5820719640c8ebe2ac6e399e8e0118fc1d5db626

Observation 229d25cc-59c3-4fb7-819e-45272197b219 · outbound

This paper cites Accessed October 22, 2023 [Online] https : / / github.com/deep-floyd/IF.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Accessed October 22, 2023 [Online] https : / / github.com/deep-floyd/IF

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.194843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:8c2a4de1bf90c25aec56a1be92f65a7055b4a1bf51f0da304ba0689eda9bf94f

Observation 7c904ef5-9956-497e-bf53-937f5145ec82 · outbound

This paper cites Accessed October 22, 2023 [Online] https: //laion.ai/blog/laion-coco/.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Accessed October 22, 2023 [Online] https: //laion.ai/blog/laion-coco/

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.198379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:0b807c4f764eacdd37772effe5b705b5f9f879f3ca0777146be8844e48fcae3f

Observation 82e01027-a6c7-492e-ae3e-d9bb8cea4c4e · outbound

This paper cites Accessed October 22, 2023 [Online] https: //github.com/hotshotco/Hotshot-XL.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Accessed October 22, 2023 [Online] https: //github.com/hotshotco/Hotshot-XL

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.201878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:4b96f24f2014fded9264312141bd3f61a777c05d814fdd8487672b851d0e3124

Observation ce0b80b9-173a-41a3-b498-75cd3b07cac5 · outbound

This paper cites Accessed October 22, 2023 [Online] https: //moonvalley.ai/.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Accessed October 22, 2023 [Online] https: //moonvalley.ai/

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.205144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:ab482f4e934ef52a3f6dac8a971310401562b20b54dcc648912c533af343e3fc

Observation d90bfc42-9362-4e03-b87f-8932679dbc4e · outbound

This paper cites Accessed October 22, 2023 [Online] https: //www.pika.art/.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Accessed October 22, 2023 [Online] https: //www.pika.art/

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.207938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:046c2a39d178beb216bc5463b73bf23261cb780ddf7bfea41380ac27bb5aa304

Observation 7f5e9669-cfa5-48a0-9c58-919dc8be5c35 · outbound

This paper cites Accessed October 22, 2023 [Online] https: //huggingface.co/cerspense/zeroscope_v2_ XL.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Accessed October 22, 2023 [Online] https: //huggingface.co/cerspense/zeroscope_v2_ XL

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.211129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:980c2332008c69a243f7a3cf42ab2ee28e579d4f93253b3518615c327efe5221

Observation 77fe6bf5-5dc0-4a8a-9b86-2f095a5ea87c · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.214729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:3001524e1bdc0618a954bd8ea723d6a9b94bd452876eb8bde6225807dfcaa1e9

Observation 17968458-04fe-42e0-841b-a467d0eca150 · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:44:22.991611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:521f32a1fec3158e98fa6df973f52344f1c56304f3d8e324633b85a63921c391

Observation 638c4a5e-76e0-4b91-b03b-161d0f927133 · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.222153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:86707d08027cfd807920320af7fd91266c82321fedc9b8a8a502ae64a38981a7

Observation a07e5ce6-1751-4c76-9b09-d8f18e2bd385 · outbound

This paper cites Muse: Text-To-Image Generation via Masked Generative Transformers.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Muse: Text-To-Image Generation via Masked Generative Transformers

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.012282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:2888d59750f69d03033cded0a6ed168948488d73e2f5e195dfc00152d761cccf

Observation 301d929a-8d94-4a41-9b87-94711272d13a · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.020027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:770ff501ab42c06fe1e15030e79ae46d6074f9edc4fbcb4d31635c1c5e819b21

Observation af3bfa9b-72f1-4c05-971a-e609f01f2c13 · outbound

This paper cites Dif- fusiondet: Diffusion model for object detection.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Dif- fusiondet: Diffusion model for object detection

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.231390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:19084aa3cd42615b330df88af11e6c7ca10d5be0984b33bf4455a1641bd585d2

Observation d7585b3c-5437-4416-932d-45b8a68dbb45 · outbound

This paper cites Reproducible scal- ing laws for contrastive language-image learning.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Reproducible scal- ing laws for contrastive language-image learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.234211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:a99c03c3e9b9c788d1292ed470a682995f55ec662c7bc9466e26012df44f5c0b

Observation 737502d4-553f-4ac1-81dc-f1cba629639f · outbound

This paper cites I2vgen-xl.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation I2vgen-xl

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.237091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:0de04917736a48be1133bd941a4bb4437b99b638f7f7e5a9f08631eadb265360

Observation 42d87797-31cd-4ee7-8f31-e653f77eaee2 · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.028439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:1c36c93314887c3a0c8eac2698a2e30b0fedb1316da08e8a9b62f45ca55caa54

Observation e83fde39-0800-44f0-a729-4a48c130bbf8 · outbound

This paper cites An image is worth 16x16 words: Trans- formers for image recognition at scale.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation An image is worth 16x16 words: Trans- formers for image recognition at scale

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.243893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:6c681540c8794d797e696e986554a7b37bbf69bcd8f1a967424235724212de83

Observation a051505e-e76b-48e6-ad71-87c41ca8ce34 · outbound

This paper cites Structure and content-guided video synthesis with diffusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Structure and content-guided video synthesis with diffusion models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.246882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:167832b7eaeb5001367cd8d25f695a8efe4423323d444c0847897a8d2889963d

Observation 984e9496-d6e7-42d0-988f-af4213b5c16d · outbound

This paper cites Make-a-scene: Scene- based text-to-image generation with human priors.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Make-a-scene: Scene- based text-to-image generation with human priors

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.250020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:58f955cea0721a96c00cff7f0acb55d10505e3e8fd22ea0047a7c69ce8b6a7d6

Observation 0570f1cd-5e16-4504-b36b-eb722b70c0ea · outbound

This paper cites Preserve your own correlation: A noise prior for video diffusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Preserve your own correlation: A noise prior for video diffusion models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.253475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:2b7e47aea0c948321198ac8396970f0ce05e12058cddb6e5539f562fe732506c

Observation 4f7a1eb0-9e3f-4007-b6df-5a23d9960ee1 · outbound

This paper cites Vec- tor quantized diffusion model for text-to-image synthesis.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Vec- tor quantized diffusion model for text-to-image synthesis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.256769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:7748001e3bf29a339c67b6321dbb12b39c9cf00230ce1731a321b03fb7483ba5

Observation 3bea67d4-04d5-42b5-8cac-6d3ed3eca702 · outbound

This paper cites Seer: Language Instructed Video Prediction with Latent Diffusion Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Seer: Language Instructed Video Prediction with Latent Diffusion Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.034480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:5cc06818d2d12798f821a59735ef0ede81057170dd8b60d6d73397f536ac1172

Observation b8b639a7-e3d9-4231-8d1a-f436b5d81943 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.040472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:8ba88590a4ee0242ac0e09d7cf83d9a8b2d1637fb58e8fd2d407d57e802658db

Observation 32643e5b-0e18-4cc9-a7f7-e855f7f6fa4f · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:27:43.523852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:5bb0b8cb852c7124575aeefeab8d2173fad106d5392c798f714c3379527bfd2f

Observation 5761198c-1171-4d06-9182-9717ae204665 · outbound

This paper cites ScaleCrafter: Tuning-free Higher-Resolution Visual Generation with Diffusion Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation ScaleCrafter: Tuning-free Higher-Resolution Visual Generation with Diffusion Models

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.052987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:65d218305d5e96e4e0a936cd48cddf024d3f271c590bf393071cb6bddaf0b5af

Observation 2a0e0e9a-d5db-4498-8b50-57b690fe616a · outbound

This paper cites Denoising diffu- sion probabilistic models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Denoising diffu- sion probabilistic models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.273435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:6b8e89f6b3f142c9aebe18b7c97f21d5c628e5d156dd5358310e536eed6c0abc

Observation 96e67082-aa48-4805-9783-b2d046023236 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Imagen Video: High Definition Video Generation with Diffusion Models

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.059650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:05e7cc43deeb9bfa9a6e664d695a899e497bed88db51e778d68c535380b2c27a

Observation 8705c65f-3e6d-4c0d-8de8-bb5ba7d82a47 · outbound

This paper cites Video dif- fusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Video dif- fusion models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.279807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:1965bc4847912cb953df303e7f57080b3db6b3ddd0f41e0c99b560be8c5fc76c

Observation a9c1627e-7a23-4c4e-8dbe-4ce0a54a5a42 · outbound

This paper cites Text2Video-Zero: Text-to-Image Diffusion Models are Zero-Shot Video Generators.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Text2Video-Zero: Text-to-Image Diffusion Models are Zero-Shot Video Generators

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.065287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:e33f9081a6295cf9a8e357e9948973d8afbbbbe333c23c2d5834a73d4462af0b

Observation 4c3dc50f-bfe1-4dbe-a7ba-5900f5fffb75 · outbound

This paper cites VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.072118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:8bcc8a953df894b7235290498da67eca1dc32efb2b0e219df659280004bd6da9

Observation 4abb260f-f8f2-45d7-bd50-eed44c057f38 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Gligen: Open-set grounded text-to-image generation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.178447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:35551c8da4987619fc043b3a8a2a3f4af0ef0015f678218eb6933d09f58fbf13

Observation 95cb688c-d060-4a28-9d31-ad61bdbb2321 · outbound

This paper cites Evalcrafter: Benchmarking and eval- uating large video generation models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Evalcrafter: Benchmarking and eval- uating large video generation models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.181696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:b4e287ac89baa03f85b00b95a78803efad817c7854893da7d9f2e4b9b57ba46d

Observation 19df5e59-6b44-49b4-8b96-eb54cc7c3671 · outbound

This paper cites Videofusion: Decomposed diffusion models for high-quality video generation.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Videofusion: Decomposed diffusion models for high-quality video generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.184806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:2da404d1d5d60c8b3631be72408c4693f9a6bb31fb4d369b185f69a56b6a2268

Observation 07386c06-b62b-4d2d-b13d-55a066e83e44 · outbound

This paper cites Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.078817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:a63b6c6b4226880a04196748535e6d03e502724292c54c9a4f2d32a504566144

Observation 567249ef-8d69-4cc6-b4b9-27a9ef9538ab · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Dreamix: Video Diffusion Models are General Video Editors

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.088460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:96037b46ada2073def6d7d6a69b945e683fdc2585c313ecc5905bdf344eac1f3

Observation 6311606b-2520-4c22-b36c-ddd467e90f8a · outbound

This paper cites T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:47:50.630300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:1a72ea222475290edfff322be3cfd53bcc5a2e854125b46a9c847a54949a2491

Observation 19294b01-280a-474a-ae30-bb0b8bbc65bb · outbound

This paper cites Diffusion in the Dark: A Diffusion Model for Low-Light Text Recognition.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Diffusion in the Dark: A Diffusion Model for Low-Light Text Recognition

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.101553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:0ac0f3c9038437651b5351091603e47a335ec0a6a634b46518cff5561368ebdf

Observation 29b7f350-b1df-4aed-beb1-a69189e5a16a · outbound

This paper cites Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.228180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:d887f8f29f40f3dad6e0e42298ee3e882138d02c715da5f4d5e1d7c537356019

Observation 9beef1a5-ce29-4775-aa10-40a30d546d3d · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.107627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:478ebe6addbfbd350ddf664a10586d1505c81359e22469ade2b275da09882432

Observation 53e65019-5e72-4751-8533-4fbfab0ceff1 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Learn- ing transferable visual models from natural language super- vision

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.260006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:633d3d4bd2a16a654be1fcbc8d370011d007446ad47bd62725c2e82b07434ef6

Observation c2d549e1-f5e3-41bc-85c9-25fa78efd9a1 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.113162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:3cb0a0a633c91b745d7d2f89b728410642c266391d8368136adb81dcf2f8b7eb

Observation 6bfdd643-8523-4fb7-8ddb-234d71c246aa · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation High-resolution image syn- thesis with latent diffusion models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.267363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:0431272f300a55bb56855e57b74c8316ef935e8412013e4577c3c580c1bfbfb6

Observation 348aacd1-35e4-4f80-887c-da449235bfc3 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Photorealistic text-to-image diffusion models with deep language understanding

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.270343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:9f72fab934768ec627cf11f4f7130504dae36f8a8c39b7898b290c8139e98ad6

Observation 9feeffd9-1cd3-4673-9890-9f2c948610b9 · outbound

This paper cites InstantBooth: Personalized Text-to-Image Generation without Test-Time Finetuning.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation InstantBooth: Personalized Text-to-Image Generation without Test-Time Finetuning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.119213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:4ab58ea89e13ee5371561b7313792e306376c3871276277ed82eec9cb4f88503

Observation 5284cea4-45bd-48b8-b6c4-909932d77378 · outbound

This paper cites Make-a-video: Text-to-video generation without text-video data.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Make-a-video: Text-to-video generation without text-video data

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.283028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:6eb4a8a39e7fb16f273ed7e443e53ef90763ea14c72712b168bdbde2b4fe3c7e

Observation afa4f77e-874d-4280-9ecd-16426018b73b · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Deep unsupervised learning using nonequilibrium thermodynamics

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.175181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:f1e7f30e7c23d373ec859a552310aeaaaa9288d2997854181c926620b0dfc721

Observation 4a85e6ce-7ec4-4cbd-9067-ced049c2a46d · outbound

This paper cites Denoising Diffusion Implicit Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Denoising Diffusion Implicit Models

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.123919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:0b9ce2bd218a9effb8edfce342043b87f27cf17ce698051f84929eef4d9666a2

Observation 94120712-75f5-436a-ae35-b25e1a5e52cc · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Score-based generative modeling through stochastic differential equa- tions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.191604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:6f0360c0d5d42bf51aa2a279ba9d90a1df1887e50eef242e57e9eaad0643c8fe

Observation ac346ada-0aec-49ff-9590-ab2cea2b113f · outbound

This paper cites Phenaki: Variable length video generation from open domain textual description.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Phenaki: Variable length video generation from open domain textual description

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.218616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:af2c86fb2d6c28461e518fc8141cbf20416ecff1edec596d6c3ff8f882483055

Observation f8a8db36-15e1-44d4-8994-62339d833ed5 · outbound

This paper cites ModelScope Text-to-Video Technical Report.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation ModelScope Text-to-Video Technical Report

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.129406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:fd68d45e266f161230eee64d4219d245f18b9122b3022bd80a6bea96dad9d261

Observation a2551e8c-de6c-4356-928a-f854648535ff · outbound

This paper cites VideoComposer: Compositional Video Synthesis with Motion Controllability.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation VideoComposer: Compositional Video Synthesis with Motion Controllability

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.135032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:25f29288d49b963689086c56ee8702527dacb639d20d06e0b081ad3548676e06

Observation 6298f9a0-5825-457e-8d03-a76f88201c28 · outbound

This paper cites LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.140329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:6ccc811b60c9e7aed50ab6947e781ec3dc67ee6a32145f2c1022d81f399fec57

Observation 7a54ab85-e063-4485-88dc-a94dfd61087d · outbound

This paper cites Make-Your-Video: Customized Video Generation Using Textual and Structural Guidance.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Make-Your-Video: Customized Video Generation Using Textual and Structural Guidance

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.145522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:9e270cf08b78433c7a0a66d0ff61de39c085309d7099aec7e23e645141bcd394

Observation 5b817ba0-aa16-402a-83f5-30185a4b7a9e · outbound

This paper cites Ip- adapter: Text compatible image prompt adapter for text-to- image diffusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Ip- adapter: Text compatible image prompt adapter for text-to- image diffusion models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.187879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:ef013d09fba250286982fcecb9ab1199c9d9e09eb296d4ada3ba92e523a15e12

Observation 01a79f2e-6c50-4810-82dc-a2c913ea487f · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.150411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:238eb8461e0b5773c08fa1d7b79794a3980318a0ad8a7df260800197a1913729

Observation f46fc3d8-e869-4386-a444-17bc8978e79a · outbound

This paper cites DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:03:58.521687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:9f96405c9d74ab9db35db6b59ee5adab71a9888cf1b8eb48c1ecd62922e50753

Observation a8853952-f1e1-4838-b391-1c392deb9ec6 · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:40:44.160267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:114234d8f7770639abed214d7605e6848ef60e2e831ca504257b48fd7155ec1d

Observation bce22b5d-18c7-4fb7-843f-aba3096c0bb8 · outbound

This paper cites Magvit: Masked generative video transformer.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Magvit: Masked generative video transformer

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.276531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:c50b60f9a5b9beadf62ac15bdd3f77bb729275117b597d7a721a7aa1160740ea

Observation eb6c908b-304e-4dc5-b1ae-b94407130d29 · outbound

This paper cites Show-1: Marrying pixel and latent diffusion models for text-to-video generation.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Show-1: Marrying pixel and latent diffusion models for text-to-video generation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.225117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:92616c719ead819130af1d91cf5cbce48cf768a1ed4899436d10fd46b381ffb6

Observation 910c30ee-a4d1-42a6-b14a-544ff9788d81 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Adding conditional control to text-to-image diffusion models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T21:40:44.240882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:198f3d5446cf99c00602896007a384d3c7e701dade736005daa04cad75e340d4

Observation 3370a45b-0ee5-45ce-a2d2-727b4f1223a5 · outbound

This paper cites ControlVideo: Training-free Controllable Text-to-Video Generation.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.165165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:baf0d54867ab3c562d8e1f329136d97de08c5b75f41b92f1f13090a59682162b

Observation 678c58a0-564f-4d41-b719-34471a499d45 · outbound

This paper cites Real-World Image Variation by Aligning Diffusion Inversion Chain.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Real-World Image Variation by Aligning Diffusion Inversion Chain

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.171638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:8d4fc60d86018c31d19eab9eb8c6aab59cb436f3c6319aa7c34b4278fa3a4e38

Observation 57894bb5-227c-4b34-b713-167d23479aef · outbound

This paper cites MagicVideo: Efficient Video Generation With Latent Diffusion Models.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation MagicVideo: Efficient Video Generation With Latent Diffusion Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:47:57.649874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:4dd3a5209d9c0da917203cce5445ddbe89a26e71e0d70bd7a71baaa273c30ddb

Pith citing papers

Observation e006bb4c-6a3b-4c02-935d-93590a2c153d · inbound

I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models cites this paper.

I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-17T23:51:24.055365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T23:51:23.877344Z digest=sha256:4172cec9d8bc26bf805b79c75d9ad91bc022b52bb956b2055f8f7a5467585753

Observation 426170da-27f0-43e9-a152-835b47dfd271 · inbound

VideoPoet: A Large Language Model for Zero-Shot Video Generation cites this paper.

VideoPoet: A Large Language Model for Zero-Shot Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-15T17:51:05.654063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T17:51:05.465548Z digest=sha256:84ee4c27a8bce6bfb5342b54947db907c896f6268976901f534337e1409f2142

Observation c1732485-e125-4bf6-80a4-b757820b66ec · inbound

CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation cites this paper.

CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-16T19:43:33.831108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T19:43:33.639639Z digest=sha256:715f9776706c2ad56a4bd6c4eb7aaf886ee94e1ab45c46cb602c3e7f8db8113d

Observation cd152788-bebd-4ade-80fc-7278477d8bc8 · inbound

OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation cites this paper.

OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T20:34:53.141898Z digest=sha256:26fb3a4466b3d5fa5c4da8ade23b8b49704bff33e6553269b552759ca73f2aa4

Observation c0b07ee2-4b25-4c54-b1b0-414e9ea2a2ff · inbound

Movie Gen: A Cast of Media Foundation Models cites this paper.

Movie Gen: A Cast of Media Foundation Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T14:16:18.521699Z digest=sha256:a68d7a6a999635e600187f42c145478a49f8170a94e29cc1697cfc8d08924b65

Observation 10460333-0b84-4b65-8e7b-13fcaeba0829 · inbound

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation cites this paper.

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-23T17:23:15.440156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T17:19:51.411937Z digest=sha256:2a9c7253ab7d139eeac63aa71c6b446b7eb1edbe54e99f69578c81dad72c0a37

Observation 5fc17b73-e264-4628-a563-8c92539999b5 · inbound

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness cites this paper.

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T18:42:02.940250Z digest=sha256:bc5443d4dbfb00395b6514b1b58b61bd80a799e7c8ac42d9b61b9c92b87c069a

Observation 380dedcd-b486-4843-bda1-495402ad3e94 · inbound

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback cites this paper.

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-22T18:56:58.362844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T18:56:39.735334Z digest=sha256:7d6cb60f9ce466b1c262e2a48b5a5518bf4e65c18fe25d930a2b8185716f7dff

Observation 5f58e02a-10c3-4ead-891b-c0498dd1c865 · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:ddad3a31dd4f11b482098114674afcc8136663d501558fd9d0c5e61c79e881b2

Observation 91015f69-6c2f-45f0-8cdc-24e0dfb6cbd2 · inbound

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation cites this paper.

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:02:10.311210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T07:59:49.398271Z digest=sha256:61202c3a30b2e5fd8fd23f1bc7ba10cafb570f54455c90f0cc66e29059350c37

Observation eb7d097e-2a0f-43da-862a-f92c445d6416 · inbound

A Survey on Vision-Language-Action Models: An Action Tokenization Perspective cites this paper.

A Survey on Vision-Language-Action Models: An Action Tokenization Perspective VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 106

Resolution
verified exact
local_arxiv, observed 2026-05-17T14:08:35.191363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T14:08:34.893876Z digest=sha256:98cf532d131feb061cdfacc95c499a422f211710bd6a2ff61ce3c93bc2303668

Observation a487e12a-7211-4226-b224-620720a0d5db · inbound

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling cites this paper.

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-19T05:17:06.610614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-19T05:13:28.767788Z digest=sha256:a0a02190a5f4189fd66db4f663ac1c408757e5776738d1f96b8dd3f74d437086

Observation 6c9fcc23-b086-4b9b-92d3-f3a8e8bcbf8c · inbound

LoViC: Efficient Long Video Generation with Context Compression cites this paper.

LoViC: Efficient Long Video Generation with Context Compression VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:21.916399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:39:21.916399Z digest=sha256:fd56516ab54d06440fa9b86ab53583f28e3c1e93b73f697eb6778b8cb1a768a7

Observation 755f72c9-b82c-4f19-a1e1-4d380e1b2978 · inbound

Pusa V1.0: Unlocking Temporal Control in Pretrained Video Diffusion Models via Vectorized Timestep Adaptation cites this paper.

Pusa V1.0: Unlocking Temporal Control in Pretrained Video Diffusion Models via Vectorized Timestep Adaptation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T15:24:25.313817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:24:25.313817Z digest=sha256:ca5285c6e9a4d51b22b8fbab5ce3bc6e21ee36111af224c95e1e3221da4c3434

Observation b4a1deb5-66ff-4975-b55a-97e8de553122 · inbound

MotionShot: Adaptive Motion Transfer across Arbitrary Objects for Text-to-Video Generation cites this paper.

MotionShot: Adaptive Motion Transfer across Arbitrary Objects for Text-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:00.075763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:20:00.075763Z digest=sha256:d6ee2987ce22c74f2b860bd84d24897d2c6663cbfbd2c8dd996d568dc8de8ffb

Observation 66558e01-953a-4bcc-91e7-70f36cf8d7ba · inbound

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation cites this paper.

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:26.298363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:26.298363Z digest=sha256:bea2c2dfa41f5f44fd3858ae84801483ece1458b856b87f5e2550d3ec4c75f43

Observation 4038557f-9098-441d-aa26-06e3cf165e1f · inbound

AnimeColor: Reference-based Animation Colorization with Diffusion Transformers cites this paper.

AnimeColor: Reference-based Animation Colorization with Diffusion Transformers VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:44.875426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:44.875426Z digest=sha256:7b72a92ef6440cca562a401cde217ef5412f0e95575f69833a8bbbd2abae80e6

Observation d12af3d5-5103-4fe5-abf5-85cedc1f9aa4 · inbound

A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea cites this paper.

A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T04:44:56.200389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:44:56.200389Z digest=sha256:3c3abc3fa47e44997748364f87537ac942c7c44159c8d37ceda828e14841a55f

Observation 0a5c21c2-a261-4a67-b674-442421b2a455 · inbound

Multi-human Interactive Talking Dataset cites this paper.

Multi-human Interactive Talking Dataset VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T04:46:04.032620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:46:04.032620Z digest=sha256:d497511f2c6f537484dfbf42685715d9d60bd0c2b9cfe8359f3106d2bc7ec33e

Observation 320308bd-f76d-4df0-b399-e7fa2cffd300 · inbound

SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models cites this paper.

SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T22:22:43.881657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:22:43.881657Z digest=sha256:dad9471c1ac7b7baab12ff39211ec49ab0d9f1b8f813c25a00ced579511874ce

Observation 67ebeb40-4eb4-4d39-b795-07e016efc5d4 · inbound

Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers cites this paper.

Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-05T22:21:05.579844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:21:05.579844Z digest=sha256:23eef5a2e935d6be12c79b53834c2b1c01e110ebae81f208cce4401869c9d367

Observation 9d89fd37-f3ec-4ff1-9b61-a42d7b042d1d · inbound

CineScale: Free Lunch in High-Resolution Cinematic Visual Generation cites this paper.

CineScale: Free Lunch in High-Resolution Cinematic Visual Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T17:46:34.279019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:46:34.279019Z digest=sha256:10a4fe63e03cd425e5002e14490bec9b067570f2eef6adb654d45ed5101e52f5

Observation 6553090b-ab34-4d09-9847-4edc1083ae44 · inbound

Human Motion Video Generation: A Survey cites this paper.

Human Motion Video Generation: A Survey VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 129

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:57.270453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:57.270453Z digest=sha256:6034c5f48cfaf6a67f0ec91eedc7d6b3d4585bd146614c5a7e99befbdc516471

Observation 8a76d11f-6c96-4ce3-96e2-ec22ffe84250 · inbound

CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion cites this paper.

CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T14:36:28.459037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T14:35:57.939589Z digest=sha256:39de6ea11c37e60d5065b052561314ae324cf2b6b1948c3535d02d7b0e184e29

Observation 51055e20-7b89-49ed-ab5a-e87a72efd481 · inbound

Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection cites this paper.

Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-04T10:54:30.288405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:54:30.288405Z digest=sha256:45fdcc7396fad6fe49d91d2acd31b61704dba15696c4a91586bd8699c0874f71

Observation 602babb6-035c-45da-8804-32abc1170651 · inbound

UniVideo: Unified Understanding, Generation, and Editing for Videos cites this paper.

UniVideo: Unified Understanding, Generation, and Editing for Videos VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T10:49:47.409544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:49:47.409544Z digest=sha256:2b0e9c5b4a2685ae88cfda5a2b6e678847d614a58228db50563f84f7fe42aaf6

Observation 7a0b5785-49dd-437e-b9e1-558f475f1187 · inbound

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning cites this paper.

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T10:45:58.066473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:45:58.066473Z digest=sha256:6d3504265c198d2c1a115385eb4a57b60e2982a905a52a9da2077450d83a11ca

Observation c7212b2e-c93a-4d14-ae43-edd1ba4f7281 · inbound

SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation cites this paper.

SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:24:18.330790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T18:21:08.172735Z digest=sha256:0b8118988e63ce6d63dda98f0a37c5dbb3a9a9271d662f8f546830b515331ca5

Observation eac69dac-033f-4718-8efa-a80d32ca4533 · inbound

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation cites this paper.

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T18:30:01.936675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:30:01.936675Z digest=sha256:d4f204f5a2b5170d4f4fd1a3e2c44116228944ad5ec309eba0765152589d7543

Observation 6dd20179-2f44-420f-b083-bc6e844aa843 · inbound

Splatent: Splatting Diffusion Latents for Novel View Synthesis cites this paper.

Splatent: Splatting Diffusion Latents for Novel View Synthesis VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:08:40.138851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T23:04:28.468206Z digest=sha256:1fdd6102d407fd8e43208d28f8be4272d37dc62b210e7cfdbfc5cb2e2ac5ec8e

Observation 1a7ad2e2-d4f0-459e-82bf-eb8759b2f0a6 · inbound

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation cites this paper.

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T13:24:43.707619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:24:43.707619Z digest=sha256:b6ac29214e219973608a6f1a9471da11dca1ccba9dcb6527d2cd9821eb9c0c10

Observation f8c335d7-9599-4d26-a8f0-d821fa852b93 · inbound

TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory Alignment cites this paper.

TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory Alignment VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T11:36:13.383486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:36:13.383486Z digest=sha256:1b5cc6ff19c1faf056400f792f5cce79cff923cfc1a771ec6c4969575f0e16fb

Observation 18a3bb33-b979-48df-b396-94677ad2290c · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:32:01.692690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T17:32:01.642256Z digest=sha256:b525e6c26bce3d643ed18ab29df3589c778e897581fb28a1a7d8098743fca759

Observation 3ac6d2d8-aee2-4119-b6ff-3e894202e66d · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-22T11:51:29.847088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T11:48:00.633421Z digest=sha256:410ef64fcf6f18a73531c805d29350bffb983c280c72071bfdb3d84ab9fa3c25

Observation 9b789e77-2c53-4e54-baac-4aa21cbaa843 · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T05:30:24.315015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:30:24.315015Z digest=sha256:1e183727cff850ed55123aef5ab919e982f830f33556fe48e723ec4ebe593a06

Observation e029e815-0cb4-4049-b07f-f681e762ba6a · inbound

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion cites this paper.

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T07:07:29.979047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T07:02:38.876518Z digest=sha256:20cc475a2a6e7f239fab548c628225f98044fd3c2f899a5bbc8a73b65c3cd99b

Observation 5f7d1632-9758-4bf8-a676-047785cf26dc · inbound

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts cites this paper.

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T00:09:55.905154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:09:55.905154Z digest=sha256:d1d808ddb258b747234c063476e8e711b4c339bc482681acf00a53954e0aad78

Observation d814426f-2eed-49f0-b1cf-764b707e8baf · inbound

EduVQA: Towards Concept-Aware Assessment of Educational AI-Generated Videos cites this paper.

EduVQA: Towards Concept-Aware Assessment of Educational AI-Generated Videos VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T11:54:09.086838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T11:52:09.262497Z digest=sha256:65138f13315f6bfa439f6c2717d2ff02678dd4770dbed0f3830f5aa936fbf8ac

Observation 6aa111e7-9027-4624-a0df-138fdfd90550 · inbound

LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion cites this paper.

LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T21:09:36.040879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T21:09:36.040879Z digest=sha256:bb641d1ccf0bdeb6fbd48c7e234485d1b55e21bd5b2b45b2d1459b39b4f9225b

Observation 4d389076-f765-4efc-9c90-ab0b0e982277 · inbound

ActionParty: Multi-Subject Action Binding in Generative Video Games cites this paper.

ActionParty: Multi-Subject Action Binding in Generative Video Games VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T00:56:26.115945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:56:26.115945Z digest=sha256:0c20df1266d8251fd5d67b3365ab78e3d37007f6a625ddb1ff7104056c2b7822

Observation 5cff4964-e024-4b99-9ff3-ca7f53f11c29 · inbound

ATSS: Detecting AI-Generated Videos via Anomalous Temporal Self-Similarity cites this paper.

ATSS: Detecting AI-Generated Videos via Anomalous Temporal Self-Similarity VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 65

Resolution
verified exact
orphan_title_repair, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T17:15:14.018979Z digest=sha256:168a3df6da32ff9bb8b05d2062ba35838573f743506239c084daf9c65fb6d638

Observation 40358177-5c14-4518-b3d0-cd0878ae0ebc · inbound

Novel View Synthesis as Video Completion cites this paper.

Novel View Synthesis as Video Completion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T18:26:39.418764Z digest=sha256:16fa207fc113567bd228a79dfaa23d2ebe44b0b4fff46e977113c759fc748a3f

Observation bd648c7c-57ac-4caa-8b8f-2894ca291ab2 · inbound

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models cites this paper.

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T18:39:26.420463Z digest=sha256:147287be127c27f5cc80d222c56c2c108e0d87a5117b6d001c007353050586b6

Observation 74486fde-550f-4174-bb44-f34241fd00c0 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T16:36:33.264166Z digest=sha256:0bc6ff55ff5306e343a53d67319270b9e97956674dc0f19f392121d4b6b1993a

Observation f23cc3f8-e384-4a5e-8ebe-e1ed8a41ded7 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 119

Resolution
unresolved
no resolver link, observed 2026-07-12T22:04:31.302192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:04:31.302192Z digest=sha256:276ade4ba55728c9aec641a380df2dfe2c78ff1a944285c858cddbe1d4c302c1

Observation 3a3761ef-a2ae-444e-a4db-5388b0a759b7 · inbound

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation cites this paper.

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:35:37.095627Z digest=sha256:a52652d51440db1d85d9f7a2ef613c6e3c9c6f9cc5056aa4df0d496db9acff01

Observation bd5d5ab6-072f-4226-9104-27b676f29936 · inbound

Generative Refinement Networks for Visual Synthesis cites this paper.

Generative Refinement Networks for Visual Synthesis VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:41:23.057949Z digest=sha256:83b4b7afb3bd6f44fc6462ce93b9b22df29dce3d0a038767a6f24c8ef82a38a3

Observation fe037970-48c3-43e6-8eca-12aa1c072dd5 · inbound

Generative Refinement Networks for Visual Synthesis cites this paper.

Generative Refinement Networks for Visual Synthesis VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T21:00:35.123495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:00:35.123495Z digest=sha256:a6f69c87f6b3cc93e56cec4da42069225e3c325f8ab8698dcfa1d785b588a0a1

Observation d607a89a-c3ce-4069-8059-afffb872a05b · inbound

CineAGI: Character-Consistent Movie Creation through LLM-Orchestrated Multi-Modal Generation and Cross-Scene Integration cites this paper.

CineAGI: Character-Consistent Movie Creation through LLM-Orchestrated Multi-Modal Generation and Cross-Scene Integration VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T05:08:37.971891Z digest=sha256:4efb6d79018259742c39a4d0d518bf102b1fae5a1c17e8757cc0e50b4c59c8eb

Observation 9600dc55-602e-435a-9ed0-b53b035d62b0 · inbound

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection cites this paper.

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T20:22:00.777817Z digest=sha256:d3b3b482fd0c2a62ccc5bbb1326086c3566ca0c389733c233f6b0157b79fac13

Observation 4d8504b3-de0d-486e-8a85-399ccb33e74d · inbound

VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector Animation cites this paper.

VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector Animation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 222

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-09T14:05:15.742176Z digest=sha256:f91c9875564513b526061ca51ffc7b55b9bb8b885431c6a028c156390eb6d06b

Observation af843475-e8fb-4cb5-8054-4bc0484bdd64 · inbound

Detecting AI-Generated Videos with Spiking Neural Networks cites this paper.

Detecting AI-Generated Videos with Spiking Neural Networks VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T16:15:42.113951Z digest=sha256:8479d90e125d5e68f21c18c1439ec70cda873223c9ab0192a4d2b64c4c28fa0f

Observation 9cd17cc0-da27-4936-be70-d0cb6ed86471 · inbound

Detecting AI-Generated Videos with Spiking Neural Networks cites this paper.

Detecting AI-Generated Videos with Spiking Neural Networks VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T02:24:57.328179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:24:57.328179Z digest=sha256:4bd7d783b9499e96ab50373769adf8d158a816da7ccce537c123c1ba3de61c97

Observation bea2588b-b945-4053-8c10-338a133b44b9 · inbound

FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction cites this paper.

FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T13:11:46.079614Z digest=sha256:83d1b1c3e7a2ad6022692c20728e362e855106e4e3afb58db8e0c8ae35b962d6

Observation 4f01954b-c2f4-4c73-b8df-fef6279dece5 · inbound

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers cites this paper.

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T01:58:25.291448Z digest=sha256:b3762e0dc0af9690ce1879bebe65b69c78ceb2034df56a75d45568bae5f01add

Observation b32a48a4-1b09-42ef-9ce0-c41cda030dca · inbound

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth cites this paper.

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T04:30:52.549395Z digest=sha256:63087757fef353e6951d3cd53a324513fd1b5c09d705b14577b51a414d5f8987

Observation 83925331-ba44-46a8-9c79-c9f184aa7f3b · inbound

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth cites this paper.

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T07:49:49.874581Z digest=sha256:f354f8121da4a34307553ea78fc25eb7a51bf15ba87ec40e5ae87c3f15e1e799

Observation d7f3fced-9737-4f04-8642-94da152fc420 · inbound

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth cites this paper.

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:27:45.071498Z digest=sha256:d07103530d22b64c2f28205b4095661a5dceb089d06a7ec87112933b7a98f5d5

Observation eb97e632-f22c-4b7a-8606-329390cd2621 · inbound

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth cites this paper.

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-20T23:03:50.742749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T23:00:25.109961Z digest=sha256:ff7e734c453afae39a457a2913f5edcde819a2bc8f0033b0c107cc684d46bd53

Observation 00242939-f173-4018-9e67-9b2ca3c5aa11 · inbound

FIS-DiT: Breaking the Few-Step Video Inference Barrier via Training-Free Frame Interleaved Sparsity cites this paper.

FIS-DiT: Breaking the Few-Step Video Inference Barrier via Training-Free Frame Interleaved Sparsity VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T07:16:33.110035Z digest=sha256:bf5dec2a38b053a0dbd47c8efe496ba650e58ef1ef7a25b9cdfe569b2e71adad

Observation daf7aee2-1696-40fa-b6fb-b7512c96db40 · inbound

Beyond Text Prompts: Visual-to-Visual Generation as A Unified Paradigm cites this paper.

Beyond Text Prompts: Visual-to-Visual Generation as A Unified Paradigm VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T06:02:39.114660Z digest=sha256:959010ca87ed0c78831ba0eab31136d7e62f8ba1b486cd1d0c3bb4de6ca69ef2

Observation 1106ef4b-2dca-4fab-b580-21bb30ce0f6f · inbound

Beyond Text Prompts: Visual-to-Visual Generation as A Unified Paradigm cites this paper.

Beyond Text Prompts: Visual-to-Visual Generation as A Unified Paradigm VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-01T14:05:46.591430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T22:18:55.933311Z digest=sha256:b8cc198e07429fc7d0d5880fb372bedd7740b265f07989db398631298b8cd9ed

Observation f544950b-2171-4303-9ae5-4d8dd2daff3d · inbound

R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow cites this paper.

R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 143

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.284627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-14T19:07:53.671769Z digest=sha256:1196a32b2fcc0750be3ab09800020c7918202f4b17c3303f8af75b4363225d35

Observation 1738c846-8730-4fd7-9bbb-4e45afe54b97 · inbound

R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow cites this paper.

R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 143

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T06:09:49.707848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-15T06:08:52.321351Z digest=sha256:b58682f8b2d765bca8a336779687760b2a8755192cf40f204b16f1be3cbe0f6a

Observation 06c10233-d274-4876-9738-623436773b8c · inbound

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity cites this paper.

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T02:33:32.135560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T02:31:48.593354Z digest=sha256:552f6cf4bb77739ae52923217e6134ed64abb6c472a522fb77d4fcdfb8ba8dc3

Observation 9c63564a-e809-4dff-aed7-53f739ae9308 · inbound

Bridging Brain and Semantics: A Hierarchical Framework for Semantically Enhanced fMRI-to-Video Reconstruction cites this paper.

Bridging Brain and Semantics: A Hierarchical Framework for Semantically Enhanced fMRI-to-Video Reconstruction VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-15T01:53:28.889898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T01:51:57.018809Z digest=sha256:764ab1db95eb2d5e85d4f5a4cd3052b36d64927761dddfa5f4677569feb5b7cd

Observation 667ae89c-62dc-41f5-9ec2-919ed76c89be · inbound

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation cites this paper.

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-20T17:48:48.649741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T17:47:32.953903Z digest=sha256:5287ae70530dda3596060db53fcae6b25cec439fbc458c385065fdab9aeadef0

Observation 21975557-6729-4c0b-9d77-5e128a21bfcf · inbound

DEVIS-GRPO: Unleashing GRPO on Dynamic Extreme View Synthesis cites this paper.

DEVIS-GRPO: Unleashing GRPO on Dynamic Extreme View Synthesis VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 44

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T20:57:47.031263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-19T20:53:42.240453Z digest=sha256:c5dc0bfbb4dbc700b666a24dae8f997e708d3cd4981ae595bc6840f6caa75d89

Observation 72c27364-3008-4e8f-873d-13a92854dac0 · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-05-20T15:08:24.927120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:6643bbcf88cefecfaaac7d5f21b263e4ebb6738705404459087fd5163f674208

Observation a413afef-44ac-45a3-95cb-e2d75af437c5 · inbound

Tiny-Engram: Trigger-Indexed Concept Tables for Generative Vision cites this paper.

Tiny-Engram: Trigger-Indexed Concept Tables for Generative Vision VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T07:49:50.273841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T07:46:05.375388Z digest=sha256:e6b358b535c6b0e577f3b581008675b3cb296c689257e69481e0aa86e0a633cf

Observation ae5d2f76-5ada-4fa2-b693-8d842953e5a2 · inbound

DTG-Restore: Training-Free Diffusion Refinement for Generative Video Super-Resolution cites this paper.

DTG-Restore: Training-Free Diffusion Refinement for Generative Video Super-Resolution VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:13:14.797079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T08:11:11.134337Z digest=sha256:ec55b407c47c826e47c87b0f538584d960e1a9973a4822fdfa0e5df8ae0c1c22

Observation 7ea8ce33-44a3-4870-87ee-b78deb340248 · inbound

CameraNoise: Enabling Faithful Camera Control in Video Diffusion through Geometry-Flow-Guided Noise Warping cites this paper.

CameraNoise: Enabling Faithful Camera Control in Video Diffusion through Geometry-Flow-Guided Noise Warping VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-28T23:22:46.711723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T23:21:12.987597Z digest=sha256:8bd03af26518bcf076f92cd8910eccfcb1fb4b15c369e5efd4f5028bcdd037f8

Observation 257bb731-5de5-4ced-977a-58182c4a0fb9 · inbound

MBench: A Comprehensive Benchmark on Memory Capability for Video World Models cites this paper.

MBench: A Comprehensive Benchmark on Memory Capability for Video World Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:22:35.146424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T19:04:58.672464Z digest=sha256:24b330a12c9b626aaae50fead91b41af31d572aba40f05ddd856a584ba71b083

Observation 5ed41403-5877-4e62-ae6b-53d61f0ef9c9 · inbound

PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion cites this paper.

PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-01T21:16:14.712435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T17:14:23.336848Z digest=sha256:e67ec391b24118c9e48c94ac48d8f7a4498612e62310b91f1ceab9370014232b

Observation 3afad763-004b-448f-a669-eb90c4f3952d · inbound

MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generation cites this paper.

MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-07-03T00:47:30.385025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T17:02:00.653422Z digest=sha256:a6bec811062495b9152ca9d0f4e92d0b4031a5e40d8ca54ae166dd581d87d365

Observation 49a0d011-ced4-4f4b-b0b4-ce06a65ec510 · inbound

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation cites this paper.

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-07-03T13:28:18.798733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T08:04:48.283908Z digest=sha256:32dab6efb684b9138dfd41c20ad8839e5e958d49966654ca4b2e3573ba9b39de

Observation e1744c67-9dbb-4be7-80a3-7d97eacb7b11 · inbound

CineOrchestra: Unified Entity-Centric Conditioning for Cinematic Video Generation cites this paper.

CineOrchestra: Unified Entity-Centric Conditioning for Cinematic Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-03T14:58:32.686951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T06:51:31.625416Z digest=sha256:0e115522cee2445c34d9ba589b9081ed8a99557646fa6f5da0d79dfc7ad3bbaf

Observation bd30e020-16c9-426d-8f0f-8163abadb71c · inbound

AoiZora: Topology-Aware Auto-Parallel Optimization for Inference of Diffusion Transformers cites this paper.

AoiZora: Topology-Aware Auto-Parallel Optimization for Inference of Diffusion Transformers VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-03T22:59:03.328953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T23:07:13.368633Z digest=sha256:f6404974bfd30831e6f7f47d7598e7a923a99c0684a8b647ff4e974c5ca7ab26

Observation fe984864-2950-4ae3-91b4-756d5ba4e34f · inbound

ChronoLock: Protecting Videos from Unauthorized Text-to-Video Personalization cites this paper.

ChronoLock: Protecting Videos from Unauthorized Text-to-Video Personalization VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-04T06:19:38.597390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T14:34:52.232360Z digest=sha256:a1df21f6760a21bdea25ec173d7024db9b9a52753500622b161d88269971d274

Observation c89c6f58-40b8-4d4e-9354-bd51e88f6a7d · inbound

Towards Error-Free Long Video Generation cites this paper.

Towards Error-Free Long Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T08:59:42.130129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T10:49:36.838492Z digest=sha256:aed1743ae713b5f0483eb6760619f2a4aaf2030eb9290f9d18e0a9e781aaabc9

Observation 035eba26-12cf-4749-b4ae-ebf98405bed6 · inbound

Ocean4D: Generative Underwater 4D Reconstruction via Medium-Aware Video Diffusion cites this paper.

Ocean4D: Generative Underwater 4D Reconstruction via Medium-Aware Video Diffusion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T09:49:44.709928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T09:24:28.715994Z digest=sha256:7422b425e6423911dc06d2e125ab725116be2b06cce0b5ec7ddb0e3810de4931

Observation 81f41235-9561-461a-af51-ccc4f3bcc340 · inbound

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation cites this paper.

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T13:19:51.145710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T05:16:53.011837Z digest=sha256:f71bdb82b954b9c45155b3a58005aa1020fead073757b071d28d01c6bc2bc7ab

Observation 0da83254-417a-4249-a41e-40b91f4e506f · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T18:25:57.831280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T02:03:45.564122Z digest=sha256:ca102ce869720adc8e2b9248f273f541ec5a41261b5d455476de15dd3d6db75f

Observation 9926009b-805e-43f8-92b8-4df82fc9d54b · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T09:35:40.495439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T06:25:58.872140Z digest=sha256:72950513cab11df85c3996f47e090112f41bc27080cfebcacd425c658a3c102b

Observation 363dcc17-b9ce-4d13-a3d7-1b7c95313dd5 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:57:22.816472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T20:52:28.444524Z digest=sha256:323c2ef57fe297eb15545a791be24989a0354f7707fa7a3c3ea9b7afcdc04b55

Observation 7a97c1de-f2bb-45fd-b4da-72fe29469e42 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T22:49:00.841691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T22:44:16.272541Z digest=sha256:543f48ae21c3331e5efaefbaafa11de666fe63f1ab1e3fced9c9238abb17bb2e

Observation cddce2ff-d97f-49b9-9f68-dd2e73ced28c · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T17:14:19.770867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:14:19.770867Z digest=sha256:9af7dc1997655c404f77f2c2bb34be458aa4a9d58a226349e428e44d01b748dc

Observation 32efcba4-0f0b-407d-b100-630e8d249af5 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T10:00:09.354702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:00:09.354702Z digest=sha256:b4f1102ce0cca26680124ffce8078ee24ab08942bf616f81230b99400c1cbbae

Observation c404dce3-3553-484a-b4d9-2948ab285222 · inbound

WarpHammer: Densifying Scene Warps with 3D Object Priors for Extreme View Synthesis cites this paper.

WarpHammer: Densifying Scene Warps with 3D Object Priors for Extreme View Synthesis VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-01T09:45:39.690419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T06:19:53.091972Z digest=sha256:dcc5c2b900bcf93050b290a640c9e5e6ab68e907577ef9eddf31776fe6ebd6bf

Observation 20b5ccad-8ddc-44f3-a09c-233c24c11e21 · inbound

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models cites this paper.

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T13:16:58.091913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T13:16:16.676244Z digest=sha256:fd9e31ed35222503f4eb2c40c499744892aea58b75a2ce9087944a605920de36

Observation f2d83883-3ddf-429e-aa5c-08cc8e59c4a8 · inbound

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation cites this paper.

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T20:58:57.242741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T20:57:06.512365Z digest=sha256:f299f7207dfc4b82fe75228b15f5d6335b14263e98b0f3e593b5744d6be1738f

Observation a750e236-be7b-481e-93b5-a9d3d8fc8632 · inbound

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation cites this paper.

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T08:45:35.247511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:45:35.247511Z digest=sha256:1f6189d713e95ef360f05a00f1d44ad09c127458dd82a559467d5c28133a63dd

Observation 5952a2ef-efa7-4dd5-8743-c68f51cbb62f · inbound

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation cites this paper.

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T02:03:40.995044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:03:40.995044Z digest=sha256:b44c4737603ba2ea75b63af7a7d23c980adff24ad6d9d4eedecd0a8f6993facf

Observation 05347fe1-8da6-4cc5-8d8f-778f307dad59 · inbound

QWERTY: Training-Free Motion Control via Query-Warped Video Diffusion Transformers cites this paper.

QWERTY: Training-Free Motion Control via Query-Warped Video Diffusion Transformers VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T16:28:38.476646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T16:19:51.348169Z digest=sha256:dc85a348fbfa7a342caf2c6ec24896cd4279e39e4a3f45f0870b944a72f0cbdc

Observation 7711f71e-9e2c-47bd-ae51-f4f156d135ac · inbound

Alignment Is All You Need For X-to-4D Generation cites this paper.

Alignment Is All You Need For X-to-4D Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-03T14:38:28.741222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T14:31:18.256171Z digest=sha256:54104a56f4340bd4fabd1c902200e753a26054cde8a7c433902785ed2e18e9e3

Observation ece77e36-243e-471c-8e05-7441f64a3e8c · inbound

Track the Noise, Move the World:3D-Grounded Motion-Consistent Noise for Controllable Video Generation cites this paper.

Track the Noise, Move the World:3D-Grounded Motion-Consistent Noise for Controllable Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T06:59:31.137097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:59:31.137097Z digest=sha256:d4530971bd1cd3ec36ee9d4f0f1f435331e96e09c7dcf1e2d56d77871a39bb57

Observation 3b447377-47d3-4821-b4de-ab659a158e85 · inbound

SPLIT: Training-Free AI-Generated and Partially Edited Video Detection via Spatial Patch-Level Incoherence and Temporal Roughness cites this paper.

SPLIT: Training-Free AI-Generated and Partially Edited Video Detection via Spatial Patch-Level Incoherence and Temporal Roughness VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T06:22:35.545340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:22:35.545340Z digest=sha256:a69389b737718eef4608de4b9568765398551c85c59fbc6380cf6e356aacc871

Observation b8824b7b-8171-458a-8764-21fe31dd8414 · inbound

ProxyUp: Training-Free Proxy-Conditioned Video Generation for Controllable Dynamics cites this paper.

ProxyUp: Training-Free Proxy-Conditioned Video Generation for Controllable Dynamics VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T00:20:07.678576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:20:07.678576Z digest=sha256:8218350ace69aa46560d68badb127b4bf7832727a064fc00db195b28169e69e4

Observation 31abe09e-a209-43be-a690-273dc9d2b82e · inbound

ShotPlan: Cinematic Video Generation with Learnable Planning Token cites this paper.

ShotPlan: Cinematic Video Generation with Learnable Planning Token VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T17:22:29.691123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:22:29.691123Z digest=sha256:4c280af206a7a014beeca8dcc2c74dff5ac2ee72d54ef3fcc03e5648e494a5e1

Observation dfb0567d-70b6-4a2a-91a3-a52d6d955092 · inbound

SphereVideo: Prototype-anchored Hyperspherical Boundary for Continual AI-generated Video Detection cites this paper.

SphereVideo: Prototype-anchored Hyperspherical Boundary for Continual AI-generated Video Detection VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:24:37.408478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:24:37.408478Z digest=sha256:768af26108c0c5f2a4f4b31a6aa4b524eaea8622a5797cfa9b5075b59bb65175