Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2305.18264.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:30.397158Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T13:45:45.958520Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 73fdab96-0d2b-4e16-bc40-035b3de460fb · inbound
VIRES: Video Instance Repainting via Sketch and Text Guided Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6487a23f-d84f-4c4d-9561-9a022dbf9b4e · inbound
Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f79c6ea-effe-4c7b-9f36-98830d650975 · inbound
Towards Precise Scaling Laws for Video Diffusion Transformers Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20f32d96-33db-4ea7-b333-4bff37593bd4 · inbound
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00df8333-4747-4c99-9041-a7187aacedb5 · inbound
Mind the Time: Temporally-Controlled Multi-Event Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b01888a-52b5-423c-bad4-5dec7185c72f · inbound
Video Diffusion Transformers are In-Context Learners Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86bebbff-4798-46bb-a95a-668f1a26e104 · inbound
Re-Attentional Controllable Video Diffusion Editing Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1ffa5b6-6822-4119-a09b-cc5bd2dd9021 · inbound
Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b99598a-eef6-4f32-9c61-f5bf50479eb1 · inbound
Ingredients: Blending Custom Photos with Video Diffusion Transformers Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb94d4e0-667d-448b-b554-ef67122ec48b · inbound
Brick-Diffusion: Generating Long Videos with Brick-to-Wall Denoising Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b72075e-1e19-4e4d-bf50-7ea9f1ed0191 · inbound
Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c201e375-546d-4903-b52e-12eca4991ba0 · inbound
Ouroboros-Diffusion: Exploring Consistent Content Generation in Tuning-free Long Video Diffusion Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5658d19e-b6df-481d-9d83-c733e44bf1db · inbound
Latent Swap Joint Diffusion for 2D Long-Form Latent Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2a4df6d-76ad-4764-9c3b-2020dd432e00 · inbound
Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 90623b86-060a-42db-ae5d-f655c1ee9ade · inbound
Long-Context Autoregressive Video Modeling with Next-Frame Prediction Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b483865c-2068-41ed-9cff-d5ffa912eca8 · inbound
Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff4c7d0-8da1-44f5-ab23-911eadd90650 · inbound
FreePCA: Integrating Consistency Information across Long-short Frames in Training-free Long Video Generation via Principal Component Analysis Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05f35cc5-2492-4ada-89cd-8018d52cd082 · inbound
Character-Centered Dialogue Generation from Scene-Level Prompts Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d5bea886-b2d0-455b-af90-1c8903bcb576 · inbound
Frame-Level Captions for Long Video Generation with Complex Multi Scenes Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b43765b-a46a-4d6e-9428-248c8e95320e · inbound
FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b31f71e-ad83-4f65-b2de-4bd54e44f6b5 · inbound
LumosFlow: Motion-Guided Long Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 914e6d65-5301-45cf-9962-ec1d1616d3f2 · inbound
Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdb76e04-ca4f-49ab-a578-ababe0539eda · inbound
Epona: Autoregressive Diffusion World Model for Autonomous Driving Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d2747e3-76a0-4025-a990-8618a86839b0 · inbound
FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9363dd6f-8a08-4e47-9055-b5e2b90c59f8 · inbound
LoViC: Efficient Long Video Generation with Context Compression Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05b365b8-e324-4737-876f-a97c2bcbaab0 · inbound
TokensGen: Harnessing Condensed Tokens for Long Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4d2d6ed-b7e5-4eba-90b7-f899695b572d · inbound
ShoulderShot: Generating Over-the-Shoulder Dialogue Videos Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f87002f-248d-49e6-a418-524b8d753cb8 · inbound
StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2667054-549a-4ea1-a644-3b69a5e24a8e · inbound
AnchorSync: Global Consistency Optimization for Long Video Editing Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d25195cc-2719-4df0-8b08-247ecd523f08 · inbound
InfinityHuman: Towards Long-Term Audio-Driven Human Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec570185-0348-40cd-a9db-958e063ce329 · inbound
RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2d75952f-1bbc-4d0a-a75e-9409d305f239 · inbound
Compositional Diffusion with Guided Search for Long-Horizon Planning Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e897bc2c-fedd-4042-a38f-b5f8dc4ae394 · inbound
FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 75a45d37-569e-4ab3-8d10-27a78811c716 · inbound
DCR: Counterfactual Attractor Guidance for Rare Compositional Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 015a5753-e019-4cc1-9b6c-81bf1ab8b586 · inbound
TIE: Time Interval Encoding for Video Generation over Events Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation af38671f-d523-4207-8a3f-f6932fd7b436 · inbound
TIE: Time Interval Encoding for Video Generation over Events Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 79e42a55-fa65-4bc7-b59c-3d1c310a73e7 · inbound
Enhancing Train-Free Infinite-Frame Generation for Consistent Long Videos Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bbe0930d-3439-4be7-8d08-c14ab80f2ca2 · inbound
DrawVideo: Generating Long Video from Storyboard Keyframe Sketches Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9b2d46f0-9827-4cb1-9f7a-9c650d46af2c · inbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3e2224ea-c61a-490d-88c8-980f2bb8fb53 · inbound
MBench: A Comprehensive Benchmark on Memory Capability for Video World Models Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c583a004-f2ae-4cee-abbc-708573ce4477 · inbound
Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9692a55f-df01-4033-9270-53051d51eabd · inbound
Test-Time Noise Guided Adaptation for Realistic Autoregressive Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8704be00-5a6b-4ac3-bb9a-8ddde7b31e4b · inbound
Self Gradient Forcing: Native Long Video Extrapolation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35d94627-f417-41eb-88f7-2719dbe550be · inbound
TPD: Temporal Prior Decoupling for Text-to-Video Diffusion Models Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.