Pith. sign in

Paper Citation Record · LEDGER

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion

As of 18 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 3 inbound Pith citation observations for arXiv:2501.05484.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05484 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:39:55.187739Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:45:59.717927Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T15:28:31.726758Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy15
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 00dfd202-73e9-4a15-b3ca-a9d37aa02890 · outbound

This paper cites MultiDiffusion: Fusing diffusion paths for controlled image generation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion MultiDiffusion: Fusing diffusion paths for controlled image generation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:56.093390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:54.954552Z digest=sha256:ddefc018e982ac5d18e3dddf23a414a424f51aa231b96f64dd8cb5f5a0250e95

Observation 6e3755cb-31e2-4f23-b470-58b5ec3c955c · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:54.963200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:54.963200Z digest=sha256:e7db5bc6c2ac3a34c7420e4640685f962d9cc7f114f396896231f1dbb376c40e

Observation 7d7713b1-3272-45f4-be30-c7ba17546dac · outbound

This paper cites Video generation models as world simulators, 2024.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Video generation models as world simulators, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:56.080063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:54.967818Z digest=sha256:b23b48e14029fb4d40a8b1083230032ee593ce55cdb596372a14551814f9666d

Observation a67a8c9d-095f-483a-b637-92638b500008 · outbound

This paper cites VideoDreamer: Customized Multi-Subject Text-to-Video Generation with Disen-Mix Finetuning on Language-Video Foundation Models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion VideoDreamer: Customized Multi-Subject Text-to-Video Generation with Disen-Mix Finetuning on Language-Video Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:54.971897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:54.971897Z digest=sha256:f9e88d99f5c3c50a930fde5e41706ad8c99814158752b8f6e8b18d811c3322ea

Observation c44e5698-a3d4-4ba2-a2f1-8aee31475318 · outbound

This paper cites ExVideo: Extending Video Diffusion Models via Parameter-Efficient Post-Tuning.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion ExVideo: Extending Video Diffusion Models via Parameter-Efficient Post-Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:54.976596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:54.976596Z digest=sha256:ea500d816920b282cedfd32a0755e70ca5f478d0af066e23e78ded2a402947da

Observation be2328aa-1086-4ec3-82fc-3d9a53b16f67 · outbound

This paper cites Structure and content-guided video synthesis with diffusion models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Structure and content-guided video synthesis with diffusion models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:56.064398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:54.981084Z digest=sha256:0247b43aa9357522bb45681e1bf7431a4cb84e9122e5901b4652e380b0a25576

Observation 0e82c70b-764c-42fb-8854-07f12dace917 · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:54.985724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:54.985724Z digest=sha256:10df7293e8cd534f318b54e9616a8b723ad4812056491af9c8b889b0897f92a2

Observation 27b5793b-627b-4dd7-a28f-5f89dedbec58 · outbound

This paper cites Animatediff: Animate your personalized text- to-image diffusion models without specific tuning.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Animatediff: Animate your personalized text- to-image diffusion models without specific tuning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:56.038920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:54.990252Z digest=sha256:8ae35de01fb9774a97eeae3ed5a5c95bd578b6ee33ef681a11cd41d072d7038c

Observation ee392231-5528-450b-92ca-1163b2bc8341 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:54.994281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:54.994281Z digest=sha256:c205eba8f2ae55ff2a2b2d5151d78f84be5322285cba91e49740220f3f9e0add

Observation bbd14ece-bef9-475b-ac73-17fb3c0dd6ef · outbound

This paper cites StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:54.999572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:54.999572Z digest=sha256:d008ed871188b9e771301b5dd7f29c63664fa247658de24e5b1428e16a8af1d2

Observation 0feaa6f8-1c2b-495f-82a4-b45739455967 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Denoising dif- fusion probabilistic models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.003977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.003977Z digest=sha256:1bb43d76df3b69ce16231ab93461018438c057e4bfc232b5a688e97baa1eefd1

Observation 90459ae5-7843-4700-9d10-68c9e0817a9c · outbound

This paper cites Video dif- fusion models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Video dif- fusion models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.008759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.008759Z digest=sha256:2294f2a5cf113bc9ed7f3fcc7aa9b6f4782c67b049a86cdd6b53939fd23510bc

Observation 2d869f6a-5867-4103-9cc5-312690d42548 · outbound

This paper cites Vbench: Comprehensive bench- mark suite for video generative models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Vbench: Comprehensive bench- mark suite for video generative models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.013503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.013503Z digest=sha256:ae44db6e4347d2bdd4bc3153ac651a81f1e673ff82c8a3178fd3c27f5b6d8439

Observation 34a6d206-968f-4167-a035-ca119d6f06cf · outbound

This paper cites FIFO-Diffusion: Generating Infinite Videos from Text without Training.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion FIFO-Diffusion: Generating Infinite Videos from Text without Training

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.018356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.018356Z digest=sha256:7e1fdcd29ddeb91a71e3f255294c18f3e3ea30e9e101b57b63c45d7268762f8f

Observation 04a02cf1-629f-4a67-98ca-f996265d0cea · outbound

This paper cites Open-sora-plan, 2024.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Open-sora-plan, 2024

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.022749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.022749Z digest=sha256:020873bc1524f1aa8b1aff96cec7d89cd46147a410832dbb76ef3287aa565380

Observation 0f7adb0b-ca44-41cd-b873-0e630534ffe6 · outbound

This paper cites Syncdiffusion: Coherent montage via synchronized joint diffusions.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Syncdiffusion: Coherent montage via synchronized joint diffusions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.975342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.026986Z digest=sha256:dbc384b3f8d33955a61961ec7b8a29dbae3c9b56cc0072cfb910961ae8de9527

Observation 3d923adf-774e-4772-8adc-8368b92ec565 · outbound

This paper cites A Survey on Long Video Generation: Challenges, Methods, and Prospects.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion A Survey on Long Video Generation: Challenges, Methods, and Prospects

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.031388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.031388Z digest=sha256:fc228266ab9b7518d3110452adcb61a5295776c48fb2eca852fb3ca821d17775

Observation 1d8717dc-d83a-4f1f-8f3b-d5350dbac0bf · outbound

This paper cites Decoupled Video Generation with Chain of Training-free Diffusion Model Experts.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Decoupled Video Generation with Chain of Training-free Diffusion Model Experts

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:39:55.490683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.035787Z digest=sha256:580a1d5375f05a10360f05a61f3cb18c966547c592b01be80a1c5d05544934ce

Observation cdd11775-0c0d-42b9-9620-cf607cd0f8f9 · outbound

This paper cites VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.040384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.040384Z digest=sha256:abe3cd3d48550e4e543ef099417961dd9693d701d7f008a2299c368b437172ab

Observation 117fbc41-a989-48b1-afc0-7a33320b9382 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.044573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.044573Z digest=sha256:6dcce745572052cff0254633cb5c79878f8ee88b87d881debc66cc33896763ac

Observation c4aef0ef-1ef5-4a05-82fc-322ee56821c7 · outbound

This paper cites VideoStudio: Generating Consistent-Content and Multi-Scene Videos.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion VideoStudio: Generating Consistent-Content and Multi-Scene Videos

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.048635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.048635Z digest=sha256:379af3681b85e903bb1220b3bae7ea5da2a944da8ffd9ed60ed5883c21d8b62c

Observation 68ade3ea-99f2-4be3-85c3-73b08d9c3dfd · outbound

This paper cites FreeLong: Training-Free Long Video Generation with SpectralBlend Temporal Attention.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion FreeLong: Training-Free Long Video Generation with SpectralBlend Temporal Attention

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.053385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.053385Z digest=sha256:0422f8c86feeee7188669eaa2404929800d74ccdb10deac27ae672195841146b

Observation 355b2bb9-29fd-487b-8e1d-1f90f045936c · outbound

This paper cites FasterCache: Training-Free Video Diffusion Model Acceleration with High Quality.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion FasterCache: Training-Free Video Diffusion Model Acceleration with High Quality

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.057965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.057965Z digest=sha256:09fb9acb0ee411c4fce4c6952dd1a22fc263c10948c74796876f3349b76aed46

Observation ec11d674-9166-4bb0-aa1c-9a2ae5997926 · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Latte: Latent Diffusion Transformer for Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.062137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.062137Z digest=sha256:2afc44778bbdddf5476261edb99856d7fc0f0ea8ed31cf912b2e223076f99717

Observation cba4747a-dae5-41d4-a980-a78ff86b5adb · outbound

This paper cites Rd- nerf: Neural robust distilled feature fields for sparse-view scene segmentation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Rd- nerf: Neural robust distilled feature fields for sparse-view scene segmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.960489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.066287Z digest=sha256:2d73da810d052f811cdfb0182a628730bf11a8894bb4c14445c7275c052eaab7

Observation fb8da623-0000-4ab2-895f-0623fa9d0aad · outbound

This paper cites Snap video: Scaled spatiotemporal transformers for text-to-video synthesis.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Snap video: Scaled spatiotemporal transformers for text-to-video synthesis

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.945963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.070123Z digest=sha256:c4d460744bc8a0ab5f22a77d539104e4804132ddf95e98299191b1f949fdd556

Observation 89b392be-3879-4d71-bad7-aee22f83fb5a · outbound

This paper cites Mevg: Multi-event video generation with text-to-video models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Mevg: Multi-event video generation with text-to-video models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.930539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.074061Z digest=sha256:63fcdf09e9cfe2898572cbeba1cf7ac6fdbfd881cc81e9c109c13e35d7ffbed0

Observation 51b58c2d-5fb9-4774-9499-6e2d48ca9a85 · outbound

This paper cites Scalable diffusion mod- els with transformers.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Scalable diffusion mod- els with transformers

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.914016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.078969Z digest=sha256:05c7cc4a14829d678349672fa5290e75715f93c6f1f8489d664a3d03e2533ab0

Observation 164acc23-3dbc-4ac9-8d28-433594637cf9 · outbound

This paper cites SDXL: Improving latent diffusion models for high-resolution image synthesis.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion SDXL: Improving latent diffusion models for high-resolution image synthesis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.896435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.083004Z digest=sha256:466235144573498d17a92a6620e2df8cb8c02a87a8051a269a4d0d928e91dfb7

Observation 6297cfa2-e6d3-4be1-bc57-d2f2f52e70a8 · outbound

This paper cites FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.087647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.087647Z digest=sha256:b1d635877b51f0bdda8a02d880c8b4f53510b4aa0bc4c4da0cb06f0907f20a31

Observation ec39f878-15e8-4212-9659-f9de5302f266 · outbound

This paper cites Freenoise: Tuning- free longer video diffusion via noise rescheduling.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Freenoise: Tuning- free longer video diffusion via noise rescheduling

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.881763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.091778Z digest=sha256:9042cd1fd3a27736827e828bffa4fd0dc62886be26d9076ae6aad7bb8a29c8f4

Observation 4020705c-14aa-41f9-a95e-8705d7948e1a · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion High-resolution image synthesis with latent diffusion models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.095664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.095664Z digest=sha256:84c71b7c988c13556a29bc1d8c35ddb02b899e0685cd2e861bbecb4b10c5694c

Observation 09bc7a3d-9814-4175-8902-fcda0ba25f41 · outbound

This paper cites Make-a-video: Text-to-video generation without text-video data.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Make-a-video: Text-to-video generation without text-video data

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.100679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.100679Z digest=sha256:3eb5f9a832245504b4c222d37e33ae9dcec694d66fffafca07a3132e8382cd5a

Observation a02ad840-4554-422e-accb-0254db5f069d · outbound

This paper cites Denois- ing diffusion implicit models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Denois- ing diffusion implicit models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.105119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.105119Z digest=sha256:e40da7f1eec35f89a5e1531f0a6016275a2fe503e66d3e867830232a3205b30e

Observation 5b72075e-1e19-4e4d-bf50-7ea9f1ed0191 · outbound

This paper cites Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.110079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.110079Z digest=sha256:99262fdabeb4d15bb10562901c3cbcd8b7f3666a7da2ec6be9b952f9eab7d3d9

Observation a32e4728-a2a7-421c-a263-96fffbe0ea26 · outbound

This paper cites LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.115023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.115023Z digest=sha256:dd1aa720f9cc3c0c275a4b26f46b5ca6d4950334451e08a5fa9be2b899d1c5a1

Observation addab3fb-7910-4c90-b151-5e3a3adc0d1d · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.119247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.119247Z digest=sha256:770173a8141bd8d26ca3c28c26aaad00d8e1a90114d609768aa27c661b4b2da6

Observation 80859059-388d-4add-84e8-eae0e44422db · outbound

This paper cites Loong: Generating Minute-level Long Videos with Autoregressive Language Models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.123934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.123934Z digest=sha256:1579ec419b9695054935bf889f029daa892618717d62b5f1492bc5f4f8dff7af

Observation 24c7f093-ee45-4784-9293-24c92276673e · outbound

This paper cites LAMP: Learn A Motion Pattern for Few-Shot-Based Video Generation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion LAMP: Learn A Motion Pattern for Few-Shot-Based Video Generation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.128303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.128303Z digest=sha256:3e1595a547aa1b434aaa56d4a4f3ed1080f37afab25045720de5eb31568bb395

Observation d39cc0cb-4fbc-480b-958c-37514faed37b · outbound

This paper cites Freeinit: Bridging initialization gap in video dif- fusion models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Freeinit: Bridging initialization gap in video dif- fusion models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.831451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.134535Z digest=sha256:480827ae69f9a8c9eca72d58797eebcf6d59bd94a3b25883aac69990e9edfeef

Observation ef14cbb0-1650-4062-8684-2ed466de0806 · outbound

This paper cites A survey on video diffusion models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion A survey on video diffusion models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.139439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.139439Z digest=sha256:7bcd8993bc27a146029b0c1edcd735f14d59edcb76ce3fa28bf1878dec12059e

Observation f813ae4e-6ced-408a-8ad1-42e5756cf694 · outbound

This paper cites TV-3DG: Mastering Text-to-3D Customized Generation with Visual Prompt.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion TV-3DG: Mastering Text-to-3D Customized Generation with Visual Prompt

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.143513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.143513Z digest=sha256:4f1fd32f080cd765a35ad3a18b41fac9df55bc213139349a6cacd69d543edc99

Observation 76ea42dd-934d-49fa-9e7f-e132fc8a856d · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.148088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.148088Z digest=sha256:be4a6c7fd800a2f12cc68afbc6449da80facd536ef1353638d719028be600d98

Observation 411bd740-6aeb-40ab-a18d-49f85d9da4be · outbound

This paper cites SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.152114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.152114Z digest=sha256:44eb8e02deecc85d4e9210cb79fe1f5760c7376d043d748401b445860142aa8d

Observation ed1d8ee9-71d4-4f7a-ab0d-200494c212a1 · outbound

This paper cites Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.156417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.156417Z digest=sha256:3194df1fc2120ffe1b4c73cb9225aaa669e84001ae716cb8d868a9976c51e355

Observation de296dab-4ec3-4eeb-9218-5115f723c022 · outbound

This paper cites TVG: A Training-free Transition Video Generation Method with Diffusion Models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion TVG: A Training-free Transition Video Generation Method with Diffusion Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.160634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.160634Z digest=sha256:2c682a5e4217a4fdd2457db5ed6c370248c9270e3336bb7f72c19ca45ababee1

Observation eacc7bf9-fe84-47c8-958d-a09b6eef261f · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Open-sora: Democratizing efficient video production for all, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.808225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.164948Z digest=sha256:8d548b7fbe5b151ae792e6475e3627c874c7c972c98eae90cd56d0fbd29ac9ce

Observation 505be5aa-083a-4ce3-938b-8cfacc59774c · outbound

This paper cites TwinDiffusion: Enhancing Coherence and Efficiency in Panoramic Image Generation with Diffusion Models.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion TwinDiffusion: Enhancing Coherence and Efficiency in Panoramic Image Generation with Diffusion Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:55.169102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:55.169102Z digest=sha256:0e06b05fe108f2a7e1e6a28f6c9f245e1bec222f56c0c1bec3dd31c1c749bbaa

Observation 0e1f697b-4a32-4edd-8d47-b7eea3cc6d4a · outbound

This paper cites First, we initialize the latent variable z′ T us- ing the Noise Reinitialization strategy, which combines lo- cal noise shuffling and frequency fusion to enhance motion diversity.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion First, we initialize the latent variable z′ T us- ing the Noise Reinitialization strategy, which combines lo- cal noise shuffling and frequency fusion to enhance motion diversity

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.777246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.178373Z digest=sha256:fd413dfc56922c8bd12c4618266c20f7fc8975a368d75595c801e4a1d9478fcd

Observation bee8b756-85aa-4a2a-93e5-64307b63c8fa · outbound

This paper cites Below, we outline the key hyperparameters and their roles, the chosen ranges for the experiments, and the correspond- ing results.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Below, we outline the key hyperparameters and their roles, the chosen ranges for the experiments, and the correspond- ing results

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:39:55.763971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.183048Z digest=sha256:369a9a60d725d5cd1eff025bd1fb823db895e4e1ad9cf13e97dbe250741f6234

Observation c5985b73-5564-46f6-912f-e723552bd998 · outbound

This paper cites an unresolved cited work.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:39:55.750221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.187739Z digest=sha256:68d7a1a6b1bb8b338ed432dd03974bbc2362eaa4d5d23ca2c330dd7efaeda68f

Observation 09f8f2f0-dbca-4583-a843-629e70c048af · outbound

This paper cites an unresolved cited work.

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:39:55.790419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T21:39:55.173886Z digest=sha256:e72097ae87701d1a5b3721df2d2729e9bcab4583d423186c082bb542456a5bf1

Pith citing papers

Observation cd3b24b7-519f-4c8c-9e91-e2a0ec074014 · inbound

TokensGen: Harnessing Condensed Tokens for Long Video Generation cites this paper.

TokensGen: Harnessing Condensed Tokens for Long Video Generation Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:28:31.802642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:28:29.815825Z digest=sha256:504142c73b7b77b0dc7d0ee4a077e54d97afe89a900011cc7e6a040b72f2c96d

Observation c0d96d2d-f9e4-43dd-b9ae-dbd997045b9b · inbound

LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion cites this paper.

LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T21:09:36.040879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T21:09:36.040879Z digest=sha256:362d479bfd9a60d19efd7f684426a9389cbb86c5cad55ee2fd4e1d10f82e1eb8

Observation a3ee96b8-e05a-4c98-add3-97d5b0cd6460 · inbound

iARCS: Iterative Agentic RL for Controllable 3D Scene Generation cites this paper.

iARCS: Iterative Agentic RL for Controllable 3D Scene Generation Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:59.717927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:45:59.717927Z digest=sha256:a610f61c258b27f2f0f57342e695863488c288f18de5d469bffb23d52317b30b