Pith. sign in

Paper Citation Record · LEDGER

Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2305.18264.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.18264 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:30.397158Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T13:45:45.958520Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 73fdab96-0d2b-4e16-bc40-035b3de460fb · inbound

VIRES: Video Instance Repainting via Sketch and Text Guided Generation cites this paper.

VIRES: Video Instance Repainting via Sketch and Text Guided Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T13:29:59.077307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:29:59.077307Z digest=sha256:659c3ec57fcf8697e00ccdf0c9733e5e859dd26c5146e0ab0686880dc15e4db3

Observation 6487a23f-d84f-4c4d-9561-9a022dbf9b4e · inbound

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing cites this paper.

Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:41.895638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:41.895638Z digest=sha256:f4d738772584f57eb4bb7b86320691c713bc8f2de033665620b0cf8dab9fa9fe

Observation 0f79c6ea-effe-4c7b-9f36-98830d650975 · inbound

Towards Precise Scaling Laws for Video Diffusion Transformers cites this paper.

Towards Precise Scaling Laws for Video Diffusion Transformers Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.299416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.299416Z digest=sha256:19bfbebaeb0ff079fa28f52d6554602d8408b031c2a856221537134a44742112

Observation 20f32d96-33db-4ea7-b333-4bff37593bd4 · inbound

Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation cites this paper.

Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T04:32:44.493222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:32:44.493222Z digest=sha256:0bd479067c5767b973d609f415d378d7dc4a26e865ef4e77c5fcb6886a34f64d

Observation 00df8333-4747-4c99-9041-a7187aacedb5 · inbound

Mind the Time: Temporally-Controlled Multi-Event Video Generation cites this paper.

Mind the Time: Temporally-Controlled Multi-Event Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T20:55:02.174263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:55:02.174263Z digest=sha256:c36bcdede842df760d027e01fcb7bf0c2df1394b8d7913c3ac0b141ae0570ff2

Observation 9b01888a-52b5-423c-bad4-5dec7185c72f · inbound

Video Diffusion Transformers are In-Context Learners cites this paper.

Video Diffusion Transformers are In-Context Learners Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T15:40:20.106573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:40:20.106573Z digest=sha256:2fa9925803873d617ddd3547c386773da882097fd0dec71c6cfeb7373f175a0f

Observation 86bebbff-4798-46bb-a95a-668f1a26e104 · inbound

Re-Attentional Controllable Video Diffusion Editing cites this paper.

Re-Attentional Controllable Video Diffusion Editing Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:09.246274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:45:09.246274Z digest=sha256:2dcbaabf894453379cc03f5067172b53aba6be999eae9531145f1608006a19ee

Observation f1ffa5b6-6822-4119-a09b-cc5bd2dd9021 · inbound

Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation cites this paper.

Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T13:15:28.534361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:15:28.534361Z digest=sha256:02bfba986e176e9261d936a9009f74d6eeb7e58bc265bca91cd2c899a0c50831

Observation 3b99598a-eef6-4f32-9c61-f5bf50479eb1 · inbound

Ingredients: Blending Custom Photos with Video Diffusion Transformers cites this paper.

Ingredients: Blending Custom Photos with Video Diffusion Transformers Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:25.918741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:25.918741Z digest=sha256:3ddf1595a60018691ae0f11ce32a4255e798c791f251f57e7f42755a9614723f

Observation eb94d4e0-667d-448b-b554-ef67122ec48b · inbound

Brick-Diffusion: Generating Long Videos with Brick-to-Wall Denoising cites this paper.

Brick-Diffusion: Generating Long Videos with Brick-to-Wall Denoising Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T22:10:11.820062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:10:11.820062Z digest=sha256:bda5b58e1a7ad5942ffca6f8a9b2fd60cfa0a362f7aab436b6dfeae615aa4b25

Observation 5b72075e-1e19-4e4d-bf50-7ea9f1ed0191 · inbound

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion cites this paper.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.110079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.110079Z digest=sha256:99262fdabeb4d15bb10562901c3cbcd8b7f3666a7da2ec6be9b952f9eab7d3d9

Observation c201e375-546d-4903-b52e-12eca4991ba0 · inbound

Ouroboros-Diffusion: Exploring Consistent Content Generation in Tuning-free Long Video Diffusion cites this paper.

Ouroboros-Diffusion: Exploring Consistent Content Generation in Tuning-free Long Video Diffusion Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T20:14:50.831933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:14:50.831933Z digest=sha256:6d6e6d1f10beab5588b0f6695915200737a48cf3d98e2f7acd943f714931eefa

Observation 5658d19e-b6df-481d-9d83-c733e44bf1db · inbound

Latent Swap Joint Diffusion for 2D Long-Form Latent Generation cites this paper.

Latent Swap Joint Diffusion for 2D Long-Form Latent Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T20:12:29.598188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:12:29.598188Z digest=sha256:325d9c9b4107527d7c1a8c314e4f6f909e72a50c3d2b3bb68da5a85f0d2b6519

Observation f2a4df6d-76ad-4764-9c3b-2020dd432e00 · inbound

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling cites this paper.

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-23T00:12:17.879039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T00:07:39.286486Z digest=sha256:3370ebcc38127f74491caabb31c5cd2d7ff569cbc47b63df18834badedb7ed56

Observation 90623b86-060a-42db-ae5d-f655c1ee9ade · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:05:17.276943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:fbcf03d32e636d6ffb5d862818ce0211eb27a9bb1e52534f4095f9beed6e16ff

Observation b483865c-2068-41ed-9cff-d5ffa912eca8 · inbound

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM cites this paper.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.397158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.397158Z digest=sha256:8d54dc0848674a93e9a48e9ebee3455c850e5a9f47b96c0c84ab9d86c43add24

Observation 2ff4c7d0-8da1-44f5-ab23-911eadd90650 · inbound

FreePCA: Integrating Consistency Information across Long-short Frames in Training-free Long Video Generation via Principal Component Analysis cites this paper.

FreePCA: Integrating Consistency Information across Long-short Frames in Training-free Long Video Generation via Principal Component Analysis Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T04:29:31.042415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:29:31.042415Z digest=sha256:64eb2c03168c3487315217b666c807695d75c0c6cd266b6895e2de9088b34b2b

Observation 05f35cc5-2492-4ada-89cd-8018d52cd082 · inbound

Character-Centered Dialogue Generation from Scene-Level Prompts cites this paper.

Character-Centered Dialogue Generation from Scene-Level Prompts Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:34:53.596659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T13:31:43.083678Z digest=sha256:84d2bdf91cdd3ec37a48fe4a9ec65f79e87a4a6c29aa4ddfcc599cc213deca18

Observation d5bea886-b2d0-455b-af90-1c8903bcb576 · inbound

Frame-Level Captions for Long Video Generation with Complex Multi Scenes cites this paper.

Frame-Level Captions for Long Video Generation with Complex Multi Scenes Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:23.659573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:23.659573Z digest=sha256:669a98dc8d14012e1b48ebfdada14702feaac31f2aaa07552921bf1c8f30a939

Observation 3b43765b-a46a-4d6e-9428-248c8e95320e · inbound

FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation cites this paper.

FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T11:54:18.828455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:54:18.828455Z digest=sha256:7feaec30b938efd980698f89a2fd99faa29f32605467871124117b740dedfb02

Observation 1b31f71e-ad83-4f65-b2de-4bd54e44f6b5 · inbound

LumosFlow: Motion-Guided Long Video Generation cites this paper.

LumosFlow: Motion-Guided Long Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:13.157728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:13.157728Z digest=sha256:56369981bec5978cdfaf7859348f25199d55082d16885a332e9fb43a2eb68c86

Observation 914e6d65-5301-45cf-9962-ec1d1616d3f2 · inbound

Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation cites this paper.

Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:59.034693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:59.034693Z digest=sha256:b3fae69a4da06503883f18c8dc3528d63f18658d5ed0db2b21425a49c85e60a4

Observation cdb76e04-ca4f-49ab-a578-ababe0539eda · inbound

Epona: Autoregressive Diffusion World Model for Autonomous Driving cites this paper.

Epona: Autoregressive Diffusion World Model for Autonomous Driving Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:29:50.767973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:29:50.767973Z digest=sha256:6cbcfe6ff0fd14dbb0bd882f092b7a1842482f289e71b9ded39b8286de58ec9f

Observation 9d2747e3-76a0-4025-a990-8618a86839b0 · inbound

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion cites this paper.

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:28:19.447426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:28:19.447426Z digest=sha256:3441d4512c23944e6e7152dc932dabd4f6f4d2bde61c00556219d6ad538f54ef

Observation 9363dd6f-8a08-4e47-9055-b5e2b90c59f8 · inbound

LoViC: Efficient Long Video Generation with Context Compression cites this paper.

LoViC: Efficient Long Video Generation with Context Compression Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:29.844730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:39:29.844730Z digest=sha256:35651b4a12bd65dad8f1f0cb96c1e4dd677e7ca15297b19c2fb051eb8087e538

Observation 05b365b8-e324-4737-876f-a97c2bcbaab0 · inbound

TokensGen: Harnessing Condensed Tokens for Long Video Generation cites this paper.

TokensGen: Harnessing Condensed Tokens for Long Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:28:30.321287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:28:30.321287Z digest=sha256:7c72e526b4d80702ad41e0eee5acd00733c6265077f2b780391215c91d2d58d0

Observation a4d2d6ed-b7e5-4eba-90b7-f899695b572d · inbound

ShoulderShot: Generating Over-the-Shoulder Dialogue Videos cites this paper.

ShoulderShot: Generating Over-the-Shoulder Dialogue Videos Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:03:24.625927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:03:24.625927Z digest=sha256:93b25c62f42fc3292d0a9a273776b503cb8fdb3ffb59c8e2c53c07ba51fd6dd9

Observation 2f87002f-248d-49e6-a418-524b8d753cb8 · inbound

StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation cites this paper.

StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T17:42:41.473093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:42:41.473093Z digest=sha256:5ce4041cf526b9a58ba4fa2143ff3f45cecf5b598eb2f53aacb627daebbbfffd

Observation d2667054-549a-4ea1-a644-3b69a5e24a8e · inbound

AnchorSync: Global Consistency Optimization for Long Video Editing cites this paper.

AnchorSync: Global Consistency Optimization for Long Video Editing Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T18:29:04.459945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:29:04.459945Z digest=sha256:e956e2c13daa42658caefdf0c5c47bbfea35b859dd9aa76da351cc5e73e3f270

Observation d25195cc-2719-4df0-8b08-247ecd523f08 · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.229734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.229734Z digest=sha256:a3c003f901df1e75dc5a2105c0d86e64c42e7e941af31d11483750b7afba2785

Observation ec570185-0348-40cd-a9db-958e063ce329 · inbound

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling cites this paper.

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:15:54.494414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T05:13:42.934115Z digest=sha256:d9879611c2758151cf7b8ed5cc4200ec6ab8dbb2b920e5f6c7d4dd9559297da6

Observation 2d75952f-1bbc-4d0a-a75e-9409d305f239 · inbound

Compositional Diffusion with Guided Search for Long-Horizon Planning cites this paper.

Compositional Diffusion with Guided Search for Long-Horizon Planning Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T13:13:47.352812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:13:47.352812Z digest=sha256:95b14ca90d9886508fa3c8343c7579e896d9993abb173fb8249e1c899479db05

Observation e897bc2c-fedd-4042-a38f-b5f8dc4ae394 · inbound

FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction cites this paper.

FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:56:07.721119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T13:11:46.079614Z digest=sha256:588bad0334e8962511ba7118200a008aeeae71b0514df81eff981e3753a2c4bd

Observation 75a45d37-569e-4ab3-8d10-27a78811c716 · inbound

DCR: Counterfactual Attractor Guidance for Rare Compositional Generation cites this paper.

DCR: Counterfactual Attractor Guidance for Rare Compositional Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:01:10.958565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T13:06:56.964670Z digest=sha256:57758dd387e3fadcb7e25b871b0d4760f6a3a3c353c36a88631a1dfa74a9a470

Observation 015a5753-e019-4cc1-9b6c-81bf1ab8b586 · inbound

TIE: Time Interval Encoding for Video Generation over Events cites this paper.

TIE: Time Interval Encoding for Video Generation over Events Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:16:28.030354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T04:26:23.558427Z digest=sha256:f06d2f238554dd80cc1fc6966f0a38ebbdae3dc57d64be2c2968c6b96dc1cf38

Observation af38671f-d523-4207-8a3f-f6932fd7b436 · inbound

TIE: Time Interval Encoding for Video Generation over Events cites this paper.

TIE: Time Interval Encoding for Video Generation over Events Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:45:45.960362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T22:47:55.768380Z digest=sha256:ad926610de25a7d349d65ecb2ae667c38c5f614d67c8dbb30b541486d4c7558b

Observation 79e42a55-fa65-4bc7-b59c-3d1c310a73e7 · inbound

Enhancing Train-Free Infinite-Frame Generation for Consistent Long Videos cites this paper.

Enhancing Train-Free Infinite-Frame Generation for Consistent Long Videos Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.904551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T10:56:16.054496Z digest=sha256:2d051ce7ff20fa2b04c16b6fb0085b7ed2acc7ba96c0a19c24afd0c16fdd7eda

Observation bbe0930d-3439-4be7-8d08-c14ab80f2ca2 · inbound

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches cites this paper.

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-25T02:40:14.549398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T02:40:00.873188Z digest=sha256:2b00cf293f603023e77d9bc85aafebbef112237c32cacba408f038a03b70b12c

Observation 9b2d46f0-9827-4cb1-9f7a-9c650d46af2c · inbound

TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation cites this paper.

TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:12:46.608468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T23:11:58.580004Z digest=sha256:739de49686bd507cfad3acf367090379daeb1b40f042cb4bb254f3dac624283b

Observation 3e2224ea-c61a-490d-88c8-980f2bb8fb53 · inbound

MBench: A Comprehensive Benchmark on Memory Capability for Video World Models cites this paper.

MBench: A Comprehensive Benchmark on Memory Capability for Video World Models Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:35.077252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T19:04:58.672464Z digest=sha256:fd7f8108540faa0c90c5fc0b199e331fe9130b84c5c0eeea5999d634a55a0681

Observation c583a004-f2ae-4cee-abbc-708573ce4477 · inbound

Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning cites this paper.

Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T20:42:37.730813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T18:24:18.903215Z digest=sha256:09c0b3614eb753e322071f3417b37677a7e4c06301ef2566dd06a421b7b13b2e

Observation 9692a55f-df01-4033-9270-53051d51eabd · inbound

Test-Time Noise Guided Adaptation for Realistic Autoregressive Video Generation cites this paper.

Test-Time Noise Guided Adaptation for Realistic Autoregressive Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T22:13:29.997190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:13:29.997190Z digest=sha256:55cc3ff5c5b058a0a109280c73261cc68f70efb429d57768a08820d4403face8

Observation 8704be00-5a6b-4ac3-bb9a-8ddde7b31e4b · inbound

Self Gradient Forcing: Native Long Video Extrapolation cites this paper.

Self Gradient Forcing: Native Long Video Extrapolation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T10:04:31.470364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:04:31.470364Z digest=sha256:723a793ec1c17352412450415f9a879322b50ff3eeecb483264d53d92a99a61d

Observation 35d94627-f417-41eb-88f7-2719dbe550be · inbound

TPD: Temporal Prior Decoupling for Text-to-Video Diffusion Models cites this paper.

TPD: Temporal Prior Decoupling for Text-to-Video Diffusion Models Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-30T23:29:03.468271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T23:29:03.468271Z digest=sha256:9792f15c46593ca317849a455d1007b786e13fe427acca6eea88021c2433ed5b