Pith. sign in

Paper Citation Record · LEDGER

Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 69 inbound Pith citation observations for arXiv:2405.04233.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.04233 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 69 of 69 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:42:54.348203Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:50:11.362384Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 18bbe2b2-819f-4c89-97c9-ff43643b7d4e · inbound

Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data cites this paper.

Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T12:06:09.126266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T12:06:09.047780Z digest=sha256:847939c8f59319b2994802ba69c747f436a87e08d573e9496029f6f422d4ee44

Observation 3f1a83bc-06f2-4c58-bf62-a40c1d175703 · inbound

Elucidating the Preconditioning in Consistency Distillation cites this paper.

Elucidating the Preconditioning in Consistency Distillation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T10:42:54.348203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:42:54.348203Z digest=sha256:1ef8fa7a2b1d7cebac5133df7d8dd2b7266f44572c5bc6f1406e446edcdf8dc2

Observation e12fadcf-51da-4a5c-b584-4ab1e0be4858 · inbound

Seeing World Dynamics in a Nutshell cites this paper.

Seeing World Dynamics in a Nutshell Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T04:42:05.761410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:42:05.761410Z digest=sha256:76e388cce6d0fcf7f3a2ce8a91db80602623d552ec9edc1eb4c002a1a705bf39

Observation b5b4c075-a60f-453f-8102-d9a4c7949249 · inbound

Goku: Flow Based Video Generative Foundation Models cites this paper.

Goku: Flow Based Video Generative Foundation Models Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T21:07:32.168829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:07:32.168829Z digest=sha256:80415174f18d751df80ec8c3229b0dc6cfea2fe4ce0b0fa269a2f85a3bea7234

Observation ff15710d-5d38-4a39-9939-51666f77cd8c · inbound

AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance cites this paper.

AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T10:09:43.574143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:09:43.574143Z digest=sha256:84f7e76a0d870d664c7ece6c0d0f408d5503c19beaf58878a4033cc43f75c2a0

Observation fc8f9a49-89b8-4ed5-918f-0115d6e9b55d · inbound

RealCam-I2V: Real-World Image-to-Video Generation with Interactive Complex Camera Control cites this paper.

RealCam-I2V: Real-World Image-to-Video Generation with Interactive Complex Camera Control Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T19:40:21.146056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:40:21.146056Z digest=sha256:d80cc2aa870807c8b8c00437bd9fcb55e6602cc3695533f66027f3f741a91acf

Observation aefb1f88-d662-4015-8ba2-8f0d9d14bed6 · inbound

Wan: Open and Advanced Large-Scale Video Generative Models cites this paper.

Wan: Open and Advanced Large-Scale Video Generative Models Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:07:14.436097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T23:05:32.595632Z digest=sha256:0f22816ee7eb25aaa03ecf7679e70d007a936ffaaeeebee9c21feff01e518242

Observation aabdfd97-d1db-47a9-a6b5-03483b016066 · inbound

MotionPro: A Precise Motion Controller for Image-to-Video Generation cites this paper.

MotionPro: A Precise Motion Controller for Image-to-Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:02.311605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:01:02.311605Z digest=sha256:538c7beb3f76335d54a92862c0ad48b462a3090d07e2e4be32dfbcbc8648943d

Observation e211f830-c275-4c09-812e-117d517cad8a · inbound

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation cites this paper.

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:28.048350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:28.048350Z digest=sha256:5ad151ebb80bde3484211240d9cfc17b3781e336e6532e49521a3126373fa668

Observation 3f76c878-9e6d-4e47-9621-6bc259ad158c · inbound

Versatile Cardiovascular Signal Generation with a Unified Diffusion Transformer cites this paper.

Versatile Cardiovascular Signal Generation with a Unified Diffusion Transformer Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:00.981355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:00.981355Z digest=sha256:b75a61c8238e5fb1aabbf9cc67c5b391e51f1fcf2416029b6f999fb280b6ae64

Observation 108ea320-d8ba-4fa5-945f-d231ff320481 · inbound

VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation cites this paper.

VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:48:03.021385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:48:03.021385Z digest=sha256:9842dc41b2c60011d7592ac61b93c4213febde3a3a792d1632660439aa60deec

Observation bbf9ac7e-42af-4848-945d-545a386a0e00 · inbound

Self-supervised ControlNet with Spatio-Temporal Mamba for Real-world Video Super-resolution cites this paper.

Self-supervised ControlNet with Spatio-Temporal Mamba for Real-world Video Super-resolution Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:18.834828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:18.834828Z digest=sha256:fbfeee734861022cf96372d5832982e3c724b935e3443028b569a6cf43b61134

Observation 5adca448-c3bb-461f-997c-2bfb5c6fbb01 · inbound

LongDWM: Cross-Granularity Distillation for Building a Long-Term Driving World Model cites this paper.

LongDWM: Cross-Granularity Distillation for Building a Long-Term Driving World Model Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:46:59.897735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:46:59.897735Z digest=sha256:1f7ea46a14c3411a1a93cb87c08e6a3a762af1f3de8c5df745e27984c4e6e389

Observation b48ef145-51f6-43df-867a-cd77217e437f · inbound

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences cites this paper.

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:55.744800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:55.744800Z digest=sha256:b0292aae12ec27c0d11d1f069ddb788b258fc5b74ba6306fecca9f234e6c9ab3

Observation 96500013-c643-49cf-bbf2-11d951ccd5f1 · inbound

Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval cites this paper.

Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:17.523287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:17.523287Z digest=sha256:d19164ff8a669f127cac5d165de3a594e87c9950d6de797c2ffd8202b2c39d13

Observation a6c3ec74-a015-42f8-9241-a2e75cc0d662 · inbound

FastInit: Fast Noise Initialization for Temporally Consistent Video Generation cites this paper.

FastInit: Fast Noise Initialization for Temporally Consistent Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:53.402851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:53.402851Z digest=sha256:cb3d062f735164227c26da83e0122fb78a0b664bf390e7cd0d88f64cbf6844c7

Observation bf5f1056-59cb-49c0-84c8-15847e188534 · inbound

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion cites this paper.

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:28:18.123749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:28:18.123749Z digest=sha256:a9358b984c26224840e3e1c06f64da47ab0cdec53d7513a089b85812a5ea04ff

Observation e7c00d3a-a334-42de-98db-1afb1b1e3d4b · inbound

Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation cites this paper.

Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:19:32.682892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:19:32.682892Z digest=sha256:48613339fbf4a96ba6e9c8e9562e15e64159903db07ef0168dc02eaf9271fb17

Observation 545dcd23-e751-45b1-8d86-22b24255c284 · inbound

AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation cites this paper.

AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:52:57.579444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T03:52:18.984005Z digest=sha256:56cb5e85f0423b048c232c9f7ef01662d9c7ecabdf957fcfe87ea6c07a5b4745

Observation 313702cf-e86a-48ed-895f-f83b9e3c3c4f · inbound

Vidar: Embodied Video Diffusion Model for Generalist Manipulation cites this paper.

Vidar: Embodied Video Diffusion Model for Generalist Manipulation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:54:28.303206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T09:54:28.271928Z digest=sha256:936db46870ed506ee4554744a0cd8972cca08eb578bca498d645bf0f5930b1ea

Observation 19fcc1b4-f978-469b-bae8-ef51b579d778 · inbound

StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation cites this paper.

StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:24.337873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:47:24.337873Z digest=sha256:96e31971c15cbbd0d2604332672133a43ca8b3faf22aba664510ebddca87d53a

Observation 36bc65c8-c9c2-4d2d-af15-bb7aad95c0f7 · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.068815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.068815Z digest=sha256:7337ab666c5318e49ef490479f89e929d8c7cd42cf02b5ac53e43b4d7d1cbe54

Observation a7ef3343-70ed-4b77-9ca3-e3a861e050d7 · inbound

Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement cites this paper.

Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T12:42:39.584121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:42:39.584121Z digest=sha256:f0092815bf2450da0bddf7d9439607788e2b571f0b40e6472886b80213029596

Observation 86905ea5-c022-43bb-8b50-c37819f87f6f · inbound

RewardDance: Reward Scaling in Visual Generation cites this paper.

RewardDance: Reward Scaling in Visual Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.396895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.396895Z digest=sha256:951e86282ec5b379474dda8e0717d992a01ee343ea4369aa56e9501394fe7671

Observation 98940530-59ab-4151-b584-309fb0b43ae5 · inbound

Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency cites this paper.

Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:51:09.187561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T08:46:16.541104Z digest=sha256:305ceff521aa0fcd8840fc2f749d38410f7b606527c01de9a622f1d9bdd29d55

Observation 5c7c8dea-5d39-4e18-bf50-10a8b8b27477 · inbound

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning cites this paper.

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T14:50:13.046825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T14:50:12.804707Z digest=sha256:8069cd202a934c04fc2606f9450876d99d96fc1ce7210cd4156e023aa0970f5d

Observation c91e4821-a5a2-4b44-be0c-2c2daf15a2e8 · inbound

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation cites this paper.

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T08:16:22.346891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:16:22.346891Z digest=sha256:cd1a7c71619f868c4a1c22186a0bf34d1dd07e4f881d4f60b3a105820a4f557c

Observation 96da0831-e0f1-4f25-ac88-8c26074f3cca · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:32:01.738632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T17:32:01.642256Z digest=sha256:b4250384205961c816f538011279d79e097fc4e8c23f94b369eff9f06b4ff517

Observation 6b6b3205-e9d1-4614-aa4b-8fcb451277c5 · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:51:29.779880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T11:48:00.633421Z digest=sha256:66ab88c044cb65451d28e9c76bfba012451b43b4acda1ff51014a89cd17974d3

Observation dc357c4b-d5fa-4e89-a476-b8612bdc3696 · inbound

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation cites this paper.

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T05:30:23.868383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:30:23.868383Z digest=sha256:329c4f9b18d1846abeb240588d014d976241fcfee406b888c18957aeb311f730

Observation 29457579-037c-4ef8-9875-4c6349a85fd6 · inbound

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation cites this paper.

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T00:03:22.328973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:03:22.328973Z digest=sha256:4a089c73886ddfe721332fa91858e8d7b4ab82da5cc6ebea5e67d5c944b2deee

Observation 6dad6221-0e31-46cf-9a7e-fa9422ad42ad · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.449120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:c5fc3089b048ad52236d74a3296687ff7ed373065a2d3007b4b5a4c6bb0aa2c1

Observation 28a8d85c-9609-4c7b-9cc7-c36e8b190dbf · inbound

RefAlign: Representation Alignment for Reference-to-Video Generation cites this paper.

RefAlign: Representation Alignment for Reference-to-Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T18:01:19.034578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:01:19.034578Z digest=sha256:9856f885ee16454ac80784b40a5204b01885ad8d15b59df8e797dbab577adcb7

Observation edd8279e-e4f4-43cb-93fa-14dc68f654bd · inbound

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation cites this paper.

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T12:23:05.876881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:23:05.876881Z digest=sha256:42817010ac8e9cf37b5b9e0b346cff17d9285282ce6c0c1414ae6013d9a857ed

Observation 906edc20-62e6-4da1-ac04-1acada01650a · inbound

ActivityForensics: A Comprehensive Benchmark for Localizing Manipulated Activity in Videos cites this paper.

ActivityForensics: A Comprehensive Benchmark for Localizing Manipulated Activity in Videos Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:00.846669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T17:06:28.335731Z digest=sha256:c56b4c9b11f9be434ab1c45681b131564f3ab472fa07e1de933fa0a72bfc53d7

Observation 820dcff3-2566-44df-90ac-38a0a4466f14 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:56.962535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:36:33.264166Z digest=sha256:fe989f4111b53ce1e78fc83ab871c465ae025c459129e06f943cbef960f81d70

Observation 99af4c32-bd3b-4f48-814e-e8f49e99c200 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 123

Resolution
unresolved
no resolver link, observed 2026-07-12T22:04:31.302192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:04:31.302192Z digest=sha256:eaffd87beac1a930d5986bc679abf9982d98a406b0dceb7895fca67c7137df33

Observation a2680bb9-fda6-47c3-b0d7-9a66a13d4cd0 · inbound

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception cites this paper.

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:26:02.297387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T14:56:00.679543Z digest=sha256:f0f181b96e86e2872b3ecd0451492278065a271bcd761df0e4dd088c62b36dfd

Observation 366aa9e4-766e-44fe-b0c0-2613841ac62d · inbound

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception cites this paper.

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T05:32:16.069091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:32:16.069091Z digest=sha256:7058d54751933b2adfdf2d65ed5a0b57cd5da4509e6fe649b5a79a8771097e54

Observation 73c73cbf-a0dc-43f8-92b5-adff13846f15 · inbound

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement cites this paper.

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.877393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:41:45.903419Z digest=sha256:11cf2388f1cf032d05a5db9d7dab915ff99b34a8f5306fcc9990180bbd16b574

Observation c9450f4f-9750-4f84-98a9-c895ccfc8914 · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:27.506068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T08:00:33.307429Z digest=sha256:bd8799796cbedf528c88a2568eccd560a72e5a05290d52657d996c9b1a7aa4ec

Observation 1ac8e706-a259-490e-ac8c-eb808e75e14f · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:14:05.854281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T09:11:02.183133Z digest=sha256:efe8ebbb921545c5761565002731b9091b97ab69f260efb1a9feb2a2aed130cf

Observation 315b9a4c-bc63-4560-9263-9f3aeb421230 · inbound

Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE cites this paper.

Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:25:47.686249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:26:58.696936Z digest=sha256:9e80fe3986ac4535ae2e5161e45839f35f0c218b69c09bd9304f095c19267a8f

Observation e00cfbf3-fd1a-4057-9b84-cc756a447b16 · inbound

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics cites this paper.

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:56:32.177838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T01:23:27.990472Z digest=sha256:026f284aca751c1d168960f155dd2d2b9973d26aceceedd6e8be8a566517098d

Observation 5fb29813-602e-4cc0-a541-577187e41b7f · inbound

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics cites this paper.

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:05:34.461197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T19:04:02.414172Z digest=sha256:808742c2a3981b00054f427a772cdb569831dd4741c45c842f82838308951d2f

Observation bdf2acaf-39ee-4f07-a0e0-1d4eb3148872 · inbound

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics cites this paper.

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:24.248845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:39:47.335022Z digest=sha256:9c045b7600e36adf92f7106e1057e6bca18e40e97c71dccc989d5488ae79fdc3

Observation 4b7632f0-21ad-41f7-a338-9b3ef276d05a · inbound

FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction cites this paper.

FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:56:07.733835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T13:11:46.079614Z digest=sha256:1143963cfbc4989c880ca88bbb996a8b30b5bb5a0f483e92a0a05426a6963de8

Observation bacb9b99-890a-47b1-9cfa-119a0473e048 · inbound

Advancing Reliable Synthetic Video Detection: Insights from the SAFE Challenge cites this paper.

Advancing Reliable Synthetic Video Detection: Insights from the SAFE Challenge Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:00:57.484674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T00:53:18.735506Z digest=sha256:f7d71c53f3ee1d6494e78ec5585357601b675038b2ab5703083d0591d722a0dc

Observation 1c0305ce-01e1-41d5-b4ce-9f1a3937770d · inbound

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers cites this paper.

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:56.284351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:58:25.291448Z digest=sha256:42624dc29bbc3f61c629f31e318ee8aafd4ade9be38abcf22b1a792153c67757

Observation 6ece8cb1-63da-427f-bf15-1e745f273244 · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:53:28.797687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T01:52:14.874049Z digest=sha256:b9d5d7db287d774ca420e220fcf7bb3380a77c9dc695b80c06da4869db3830a2

Observation e0883180-b879-40c3-a4f4-a89f27e48976 · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:59:06.448785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T21:54:33.902256Z digest=sha256:22e7244462716c933931e8becd36ec58bbce0ca73a01ca68af0affcdca036f12

Observation d30ec527-f7ec-401b-a07e-30a280e6b492 · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:14:05.673014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T09:12:35.777810Z digest=sha256:72a7337fb59cf07ae19b529bde773106d36b8e88449857bbc80c3ab84b9fcdbb

Observation 276a741a-360a-44a7-a800-be7d3fe1e9f0 · inbound

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation cites this paper.

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:45:05.499291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T21:44:21.427570Z digest=sha256:5c504dcdb8d57b1bb85acfbb221c1eb033d3406e5783fdea048d97ee65df06ca

Observation 5d5d664b-2c8d-4fa7-b654-8779667e78a5 · inbound

Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation cites this paper.

Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:05:04.406243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T20:59:34.847496Z digest=sha256:686fd0d343b72810a2b65c27939d8f47e08fb39023eb01c1ac010dc4c025d62d

Observation caa404be-6042-495a-af23-765adadb3387 · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 150

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:08:24.963760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:cb34d28fa51a033d1c9715da4b6fb35566460e4b366a4f398983f11c2817ae17

Observation 1d7cd6f7-7a09-469a-ae35-a5751b533910 · inbound

minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models cites this paper.

minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:23:14.795387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:22:48.074724Z digest=sha256:d3f0cd787287359d0d9ddbf7eafb485a312499447da82af80a7869d2a2f21575

Observation 56fda64e-c019-4d5e-b6af-9dc56e19871c · inbound

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation cites this paper.

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:12:25.196547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T17:05:57.685728Z digest=sha256:0c86d0bcf23ddac59ca7d83bea16d2dcd787f1cdc77437210bb96ec954dede94

Observation 46c5aad6-4827-4c51-9342-8c0eab89d6ad · inbound

Explainable Forensics of Manipulated Segments in Untrimmed Long Videos cites this paper.

Explainable Forensics of Manipulated Segments in Untrimmed Long Videos Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:56:20.987409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T14:49:11.015838Z digest=sha256:1c70823a513d9c9795d042106fa9f774fff457e32c4386a4235beaeed432b7b3

Observation b72fe840-58c4-4a2e-a7ce-0ada4fa6ed0b · inbound

BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression cites this paper.

BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T11:59:12.867166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:59:12.867166Z digest=sha256:6e1a35361eec38a17525f8b0d6e2f9a6ce249530d6866a9afd14778318db52b2

Observation 3f0cf397-87ba-478d-ad4a-7d92ecddd226 · inbound

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation cites this paper.

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-03T08:17:45.973433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:46:56.871174Z digest=sha256:331670907326fef152bcf12d7f076054ca8a31870ff0279a88a629c3cfe6628c

Observation ca3c9ce2-55ce-4f24-8cad-efdb17ea1565 · inbound

Pulse: Training Acceleration for Large Diffusion Models with Automatic Pipeline Parallelism cites this paper.

Pulse: Training Acceleration for Large Diffusion Models with Automatic Pipeline Parallelism Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:39:24.664412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T19:36:14.248003Z digest=sha256:3f23660af14cecbb2fdb4f0d75213e44c778ea7509f7c67ca81e39c2be9ba782

Observation 8070024f-e6c1-40b7-a7e0-4f4c5f45b318 · inbound

GroundShot: Visually Consistent Multi-Shot Long Video Generation via Entity-Grounded Shot Scheduling cites this paper.

GroundShot: Visually Consistent Multi-Shot Long Video Generation via Entity-Grounded Shot Scheduling Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:19:30.253059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T18:15:51.862192Z digest=sha256:9935e72c0243b92b74d0caa41dd745cf2a7f9f9f2a8455942264b224f8ead3c6

Observation 63236064-2202-4cfd-b6a6-927e9602c6e8 · inbound

GroundShot: Visually Consistent Multi-Shot Long Video Generation via Entity-Grounded Shot Scheduling cites this paper.

GroundShot: Visually Consistent Multi-Shot Long Video Generation via Entity-Grounded Shot Scheduling Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T10:48:04.099603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:48:04.099603Z digest=sha256:441d2eede87b1af87017e91c97ef638e76e97ef0f309bc515fb79e05e71d4486

Observation f1e06150-11b4-4595-9c6d-9954ef5f4660 · inbound

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation cites this paper.

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:20:05.888697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T21:33:38.643889Z digest=sha256:11734469947728b5c11ebe22e10623785d4d0fc0967fd5f0989e75bfe5577278

Observation bfb82141-da1c-4737-86d2-ac420b481bab · inbound

Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models cites this paper.

Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:50:11.364149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T20:57:30.765802Z digest=sha256:cc293e0c761a91fc4538e267dff15807c76311f97b07e92ec8ddee15d2d54d2b

Observation 23814f50-bb84-46bd-bfab-beb691e994f2 · inbound

MemLearner: Learning to Query Context memory for Video World Models cites this paper.

MemLearner: Learning to Query Context memory for Video World Models Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.915950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:07d2e1954adf003c7f93b7296cc773ae47d8becc3061067bf1bf7b881c52248b

Observation f9db2bae-7545-4317-a7c5-37aff2e01a3e · inbound

MultiRef-Compass: Towards Comprehensive Evaluation of Multi-Reference-to-Audio-Video Generation cites this paper.

MultiRef-Compass: Towards Comprehensive Evaluation of Multi-Reference-to-Audio-Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T03:12:33.331799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:12:33.331799Z digest=sha256:cd6a5c13c15ac48aad24eb78de9fa0ddca615f5837579f6ab40f8befe0e04ad3

Observation fec10dbc-a174-4198-95bb-6b713a840974 · inbound

FilmBench: A Film-Grade Benchmark for Cinematic Video Generation cites this paper.

FilmBench: A Film-Grade Benchmark for Cinematic Video Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T20:13:31.342466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T20:13:31.342466Z digest=sha256:de27761b4b31886a29dfdfe0d273d48fe750d63bf8f08d708ad4e2d63573ad37

Observation 8c14f1c3-cab0-45e0-b77e-7aa0ca7795d9 · inbound

WorldExam: Benchmarking World Models from Apparent Appearance to Inherent Reactivity cites this paper.

WorldExam: Benchmarking World Models from Apparent Appearance to Inherent Reactivity Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T03:23:53.366492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:23:53.366492Z digest=sha256:a33b53b2717ed6ef08e91f181a27354fba0227abc654b894622b1f4d271d7d3d