Pith. sign in

Paper Citation Record · LEDGER

Towards Chunk-Wise Generation for Long Videos

As of 12 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2411.18668.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.18668 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:11:52.136614Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6deccf3c-9644-4f84-8d6c-7246b6df6751 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Towards Chunk-Wise Generation for Long Videos Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.035699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.035699Z digest=sha256:536b02205da129e091ece1db2c3573294b34b9acae6a6ea10b62ca4d8dbfeaaf

Observation d2adc6c2-350f-4cf7-b45a-19b43094854b · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Towards Chunk-Wise Generation for Long Videos Emerg- ing properties in self-supervised vision transformers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.040300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.040300Z digest=sha256:ab50edfcafbe492f0f34745eead36984ee71e492454ff4a6d68fdc5791dc47a8

Observation ff646f98-332d-4a4c-8793-f853c87eec05 · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffu- sion models.

Towards Chunk-Wise Generation for Long Videos Videocrafter2: Overcoming data limitations for high-quality video diffu- sion models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.043918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.043918Z digest=sha256:3eda9ab4cd6be37bed7f0278be2d81830288a1bddc0043e6b8e224233e43f659

Observation bc0ed4ac-52f3-495d-b810-f86a23f3bde9 · outbound

This paper cites Control-a- video: Controllable text-to-video diffusion models with mo- tion prior and reward feedback learning, 2024.

Towards Chunk-Wise Generation for Long Videos Control-a- video: Controllable text-to-video diffusion models with mo- tion prior and reward feedback learning, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.381472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.047471Z digest=sha256:01e896094ebbaa9f8103c5dc9288897b8a2ca731a5ebe933e25996bd978711f3

Observation b1dd4b0d-9967-4a93-9e36-696aa6d95008 · outbound

This paper cites Generative adversarial nets.

Towards Chunk-Wise Generation for Long Videos Generative adversarial nets

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.370925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.050893Z digest=sha256:022878c0cacb7f70b20c4de5bafe5d0b84e7dfb78f5dad1e54235aaa84c794c2

Observation 9de4b07a-3c29-4c95-a8c1-e384d9385910 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Towards Chunk-Wise Generation for Long Videos AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.054677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.054677Z digest=sha256:56eff6776ee68def9ca9d9c87d212b49b196d8f8bfc7d7ab45f56d320a3432e3

Observation ef0e1c46-7740-42bf-8405-1ccd95fd6cac · outbound

This paper cites Denoising dif- fusion probabilistic models.

Towards Chunk-Wise Generation for Long Videos Denoising dif- fusion probabilistic models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.361551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.058474Z digest=sha256:8a99e977584388df9dbb03a6e2f5b275ab537b2378a688ef49bbaca8601df0e9

Observation 895c75ae-0d8c-41c1-81ac-8bb628799d9a · outbound

This paper cites Vbench: Comprehensive bench- mark suite for video generative models.

Towards Chunk-Wise Generation for Long Videos Vbench: Comprehensive bench- mark suite for video generative models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.061484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.061484Z digest=sha256:50171508be5c10347434d56871c62200a1ef31de34c8abc32b80bf903294e2ea

Observation 33a7be06-7b5e-4eb1-bfb5-fae8ede5089e · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

Towards Chunk-Wise Generation for Long Videos Elucidating the design space of diffusion-based generative models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.346238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.064876Z digest=sha256:5a8cc901bfa3bd9b3632743018a4d0e42a967caa43e07676299b8a64afc1b968

Observation 773a7575-2607-48b2-b67c-115f1c664f0c · outbound

This paper cites FIFO-Diffusion: Generating Infinite Videos from Text without Training.

Towards Chunk-Wise Generation for Long Videos FIFO-Diffusion: Generating Infinite Videos from Text without Training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.068105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.068105Z digest=sha256:c18a2d1dd558f5780ba348cd12a4b10274bceac9ec00c65a8d7d252d52123e4a

Observation 0dc10f3c-39fc-4eb2-b484-f002e9a88bcf · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

Towards Chunk-Wise Generation for Long Videos VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.072084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.072084Z digest=sha256:7fb3ffa9a842756b14a86a6417ca6408a2ec2cbb8d5eda3d334e6a61ea6a625c

Observation f97c8577-1f85-4e29-93db-3b1197e4681f · outbound

This paper cites Open-sora-plan, 2024.

Towards Chunk-Wise Generation for Long Videos Open-sora-plan, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.335102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.075833Z digest=sha256:6adf57f2f4c39cfe41b62a8212ab9fbf5f435bb692ed2318b4d817f1355450eb

Observation c2e80aca-341b-472d-9b7f-e1fee4cd7279 · outbound

This paper cites aesthetic-predictor.

Towards Chunk-Wise Generation for Long Videos aesthetic-predictor

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.325605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.078862Z digest=sha256:fc25086885dfd3f0c68c60be46a9169385d2621271bf36c564d9e58cc5defb9f

Observation d18bfcda-a372-4a92-8923-cf0bc47cf535 · outbound

This paper cites Amt: All-pairs multi-field transforms for efficient frame interpolation.

Towards Chunk-Wise Generation for Long Videos Amt: All-pairs multi-field transforms for efficient frame interpolation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.315535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.082099Z digest=sha256:4e3cda7f9d1731fb8c3d946497ed0dc2f822740e8ebfda2efcf322650cbd0bf3

Observation 3cda880c-f228-42f2-8032-ac04a190fbc5 · outbound

This paper cites Pseudo Numerical Methods for Diffusion Models on Manifolds.

Towards Chunk-Wise Generation for Long Videos Pseudo Numerical Methods for Diffusion Models on Manifolds

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.085150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.085150Z digest=sha256:9d85dabbfda8e9870bb484dbcb75752cf5c37df8de2b86fe7f321c1e8c9cbe46

Observation 9327c4d9-2c11-4f41-9001-a7787ab9e142 · outbound

This paper cites Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.

Towards Chunk-Wise Generation for Long Videos Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.088566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.088566Z digest=sha256:731618f45ce2c43708bcb644045538be556dba0bd928f5dd67dd708df8262edc

Observation e1c6c4fa-c7c1-4d74-875a-5fd72fa8fdb9 · outbound

This paper cites Scalable diffusion models with transformers.

Towards Chunk-Wise Generation for Long Videos Scalable diffusion models with transformers

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.092574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.092574Z digest=sha256:ceb8e43efadb494fe1d4ab10ffdf604e9aa3f931a8890efa6d1b5701103714d8

Observation 5a06b7bc-f90e-4989-9590-ab77a60d5514 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Towards Chunk-Wise Generation for Long Videos Learning transferable visual models from natural language supervi- sion

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.096109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.096109Z digest=sha256:347f6be472d96540b413030034bb1c3bda28bc14cae15c30d2ae91b46ab8fff0

Observation 628339fd-d702-4bbc-9bc4-b73dbda6c8d3 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

Towards Chunk-Wise Generation for Long Videos Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.099505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.099505Z digest=sha256:a1cd61e17c7e5feab6d822dcad1b1ed28efb6f9ed8d205a5ddf7b26879d148b5

Observation 7a508ef1-e9d1-41a8-8578-68e9141c205f · outbound

This paper cites ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation.

Towards Chunk-Wise Generation for Long Videos ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.102699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.102699Z digest=sha256:98646a7688361bfa30b25a640b5c50ad8e95801f87898d45d956df5f0f42cfa2

Observation 8aa0ef68-bdc1-41ce-947f-08be4808a00f · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Towards Chunk-Wise Generation for Long Videos High-resolution image synthesis with latent diffusion models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.106213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.106213Z digest=sha256:11953f85255781310b281e720c6777018d8471f3e2c23631bb636ae89de93b1a

Observation 29ad6fe1-c1e2-40c2-93b5-b31a21450ee9 · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

Towards Chunk-Wise Generation for Long Videos U- net: Convolutional networks for biomedical image segmen- tation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.109996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.109996Z digest=sha256:41d92414721415ae95bd79646fafb223a4a30d497acb7569f335378d91f04625

Observation 8a3b6530-ecb4-45d7-b5e0-a4de2393a5a0 · outbound

This paper cites Denoising Diffusion Implicit Models.

Towards Chunk-Wise Generation for Long Videos Denoising Diffusion Implicit Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.113092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.113092Z digest=sha256:a97c472aa0a254ec129e5e49983ec9854822675fe085f4026ab425ade75801ff

Observation 6f113df8-9ceb-412a-bf98-5e03a32bc4ce · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

Towards Chunk-Wise Generation for Long Videos Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.116505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.116505Z digest=sha256:7092618ce2bb7117c9099677711f27c9c9c540b09b04f88f7dffe48035fa67f8

Observation f95ad6d7-4238-493b-a81e-e6cfcfefc142 · outbound

This paper cites ModelScope Text-to-Video Technical Report.

Towards Chunk-Wise Generation for Long Videos ModelScope Text-to-Video Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.120124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.120124Z digest=sha256:6b2fa1ec4579cdf1b57411f84126015d278b67619ee91d08c8e8ebeb732a48ff

Observation b605e088-94a8-4fff-9b43-0cfd0b6ac3b7 · outbound

This paper cites NUWA-Infinity: Autoregressive over Autoregressive Generation for Infinite Visual Synthesis.

Towards Chunk-Wise Generation for Long Videos NUWA-Infinity: Autoregressive over Autoregressive Generation for Infinite Visual Synthesis

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.123640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.123640Z digest=sha256:75e1c6dee34935cdf9688f7633c421c3dc812a3c110e438c737225b8c465e02d

Observation 73b0e0a6-9475-492b-a316-e2de0df150e6 · outbound

This paper cites Freeinit: Bridging initialization gap in video dif- fusion models.

Towards Chunk-Wise Generation for Long Videos Freeinit: Bridging initialization gap in video dif- fusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.273333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.126830Z digest=sha256:521d94edd18300348d63fec2cdceaec7a4d77b99f39cb0aeacba1d2c2b87af15

Observation bac21d8c-b2d1-4039-bb4b-a8c678fc9ac0 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Towards Chunk-Wise Generation for Long Videos CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.130074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.130074Z digest=sha256:bb4d6ba95a3a43906ef2ed3dea41d16248d18cd8acdd2122d4ffc25f718dc192

Observation efceeb06-a411-4c85-ab80-80fa2faac988 · outbound

This paper cites Magvit: Masked generative video transformer.

Towards Chunk-Wise Generation for Long Videos Magvit: Masked generative video transformer

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:11:52.262280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:11:52.133459Z digest=sha256:5a155751ebc3bfe752ad21f817bf234588632c94f4adc336838eb70c0ae24ce9

Observation afdcb13c-091e-438a-b927-0962ed38e2d7 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

Towards Chunk-Wise Generation for Long Videos Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T11:11:52.136614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:11:52.136614Z digest=sha256:a8dc7b97bcf3677c98a39c94258b12bf3b21658cc5c41aef2e298843ddf5bfb6

Pith citing papers

No inbound Pith citation observations are available.