Pith. sign in

Paper Citation Record · LEDGER

Long Context Tuning for Video Generation

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2503.10589.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.10589 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:39:23.564667Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:59:42.126868Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dd22fff3-da1f-4b98-b803-84bfa263a91e · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction Long Context Tuning for Video Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T23:05:17.409270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:7f0792c798cdfe37c46878591ee6d4bba2224569b2a93a0725ce59de4ad06949

Observation 6c0d6521-52c3-41c3-ad36-e9ab732fda96 · inbound

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion cites this paper.

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion Long Context Tuning for Video Generation

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T01:36:53.332342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T01:36:53.029590Z digest=sha256:fe0f51c5f76bbf8e095f51bf057f9cb74f2ed6a7fafb588d3d33840824f38b25

Observation f7dcf032-3a38-4c2d-ac4b-43df9576d2bc · inbound

Seedance 1.0: Exploring the Boundaries of Video Generation Models cites this paper.

Seedance 1.0: Exploring the Boundaries of Video Generation Models Long Context Tuning for Video Generation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:09:57.472650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T12:09:56.836351Z digest=sha256:c6f7c5742ce71b280aaf88c58903aaa0120576a51786ddb78429378084d9dc5a

Observation 02620058-ac81-4ad0-96f2-5e083b777a16 · inbound

LoViC: Efficient Long Video Generation with Context Compression cites this paper.

LoViC: Efficient Long Video Generation with Context Compression Long Context Tuning for Video Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:23.564667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:39:23.564667Z digest=sha256:5c335018c4674ac4cda457690b7f4315e2301308140cd62eedb376025e71fdfd

Observation 42fb464e-0f95-46f8-bc8d-19593f606416 · inbound

Captain Cinema: Towards Short Movie Generation cites this paper.

Captain Cinema: Towards Short Movie Generation Long Context Tuning for Video Generation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T14:37:19.522905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:37:19.522905Z digest=sha256:21d0f2a06325a20f6f13ff07ed276bd4f5dcfca28eb8c9fa27ec426f1d5c98aa

Observation 1cc8a74b-56f0-4691-af30-bc092bd97b36 · inbound

Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation cites this paper.

Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation Long Context Tuning for Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:20:21.169944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:20:21.169944Z digest=sha256:403c8efb8becdaa96729ec8b12c884fe80d8d5fa3c544bc2ca7978807683a548

Observation 5c5c933a-c49c-4ec7-b88b-ea46b8f5b030 · inbound

From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms cites this paper.

From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms Long Context Tuning for Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:17:45.932208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:17:45.932208Z digest=sha256:51e2762fc429a42290833417aa45919198668322ccd456ccaaeed1dc2a529bf6

Observation a546563a-4eeb-49d1-8ed4-2e16c6914f7e · inbound

LongLive: Real-time Interactive Long Video Generation cites this paper.

LongLive: Real-time Interactive Long Video Generation Long Context Tuning for Video Generation

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T03:52:59.389278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-15T03:52:59.287555Z digest=sha256:37dd17c283b891eafcab764464dcaf1651427f5ed8fa2ac5efc442aebf9ea703

Observation fef2babe-f1e1-44b7-b827-d25d7f1ac612 · inbound

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time cites this paper.

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time Long Context Tuning for Video Generation

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:15:29.182969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-16T11:15:29.102090Z digest=sha256:494a2473881ccbf169e774489c47a11f520d810c353291c7becdf5a712364231

Observation ae8f9311-cb65-4e1d-bf2e-b5beedf41576 · inbound

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning cites this paper.

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning Long Context Tuning for Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T10:45:58.380360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:45:58.380360Z digest=sha256:e35ad1bc12c1778aae4fc0a692946a330e14544daeccd6bd29e78be1b6dad047

Observation d9689b37-09f8-4110-aae7-52c5025e35a6 · inbound

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation cites this paper.

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation Long Context Tuning for Video Generation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T05:19:04.994702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T05:14:36.280155Z digest=sha256:a0f0d50572f51f18a9a56612eb7024627c4a669f0811fe72cf7c15c4eb57f640

Observation 1eb05b3e-3471-4a3d-a8bb-9429329417c6 · inbound

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation cites this paper.

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation Long Context Tuning for Video Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T19:41:56.447590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:41:56.447590Z digest=sha256:d48c8d878d6c70ab91557aac7089fadeb4714b46c71fb22887e16fe0e06e46a3

Observation deebd76c-ea70-4052-b8a0-fd6cc89530be · inbound

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation cites this paper.

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation Long Context Tuning for Video Generation

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T18:17:55.154034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T18:17:54.943863Z digest=sha256:ccb9b73b2073d3cdeb0e42a88f81ec5002af525ba3ac91352785521671ce80ef

Observation a33be3f4-e2c9-4d37-9190-3331c9786a2f · inbound

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models cites this paper.

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models Long Context Tuning for Video Generation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T23:01:20.365367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T23:01:13.910539Z digest=sha256:24ed3835daad3fa2cf1c068a24637fb390d9d9f5502423249bbcfbe646538801

Observation 9dd4f7be-ebf0-4929-82f6-4fd78bd08520 · inbound

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling cites this paper.

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling Long Context Tuning for Video Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T15:46:07.570758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:46:07.570758Z digest=sha256:ef1c0ed98e8ad3128cde5f6128701af0b7ef9d4b988cd7cddbef46a377f4c992

Observation 3ca36f0d-2cba-4e07-ac0b-66ac5cafc9d4 · inbound

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion cites this paper.

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion Long Context Tuning for Video Generation

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:07:29.956398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T07:02:38.876518Z digest=sha256:2be1db7ee6e0847a79a4d7cc26679c69cb929368c1876f98a3bf2f488fa17d12

Observation c47f3ddb-b99b-4962-a57b-95317005655a · inbound

Towards Sparse Video Understanding and Reasoning cites this paper.

Towards Sparse Video Understanding and Reasoning Long Context Tuning for Video Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T23:33:09.853077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:33:09.853077Z digest=sha256:e9ab3fa6cd7bc1ae9bb060c48aadcf69afb62a4645c9754c495f76cc6f0f30de

Observation 509a244f-c2d5-4f83-971f-afc8e854ec17 · inbound

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation cites this paper.

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation Long Context Tuning for Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T12:23:05.876881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:23:05.876881Z digest=sha256:66e3076507783ff91da5a0b94933909dd2af8adab4234f513ef199a3c2aca9db

Observation e384bf9a-0c13-4b1e-bd13-51e1e9142af0 · inbound

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling cites this paper.

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling Long Context Tuning for Video Generation

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:25:53.467071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T19:08:56.588282Z digest=sha256:88990d05e6b5cacf7b93a088d13d4609b2bb3ea737246a606d1343459ba0789b

Observation d36582f7-8dd8-4c34-8da9-4c92f35ce56e · inbound

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding cites this paper.

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding Long Context Tuning for Video Generation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:11:03.849772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:07:45.595260Z digest=sha256:5b814b253ac56e5f7a129aefe83f3ff75b27e8ccce7b4cbe718ff49ab6f5ebb8

Observation 2b2887c9-0918-4e57-a3ca-bfbdf030485a · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges Long Context Tuning for Video Generation

Reference 284

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:26.131736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:1cf2b49e654039381cc66a1c93723719caee08f84dcfaad5ca91d87a979bec26

Observation 3516844a-2748-4f4a-8241-3d94b56d727e · inbound

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity cites this paper.

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity Long Context Tuning for Video Generation

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:33:32.195452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T02:31:48.593354Z digest=sha256:abeac330501a4e562ce852ffe53b74f229096d4dd5d60d8c7e6fe4134bfaa123

Observation f29817ed-b14d-4853-a33e-34511bee3852 · inbound

EM-Vid: Training-Free Entity-Centric Memory for Efficient and Consistent Multi-Shot Video Generation cites this paper.

EM-Vid: Training-Free Entity-Centric Memory for Efficient and Consistent Multi-Shot Video Generation Long Context Tuning for Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:22.344529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T04:57:13.479165Z digest=sha256:232f0c97ebf544fc231840fb1ac5ba85af24868c36c052c271f842ac6ec048b6

Observation 240f8d65-c023-4d2d-8fb0-7737c9999ac9 · inbound

DisCo: World Models with Discrete Camera Motion Control cites this paper.

DisCo: World Models with Discrete Camera Motion Control Long Context Tuning for Video Generation

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:37:22.470643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T20:21:20.257532Z digest=sha256:3cfdb50ce9de73b91b8990934d50639ef9bdeb44bbc224f0d6282fa78d13d005

Observation 073bfb7a-b621-4518-9f95-be6ef80f2a32 · inbound

UnityShots: Memory-Driven Multi-Shot Audio-Video Generation with Boundary-Aware Gating cites this paper.

UnityShots: Memory-Driven Multi-Shot Audio-Video Generation with Boundary-Aware Gating Long Context Tuning for Video Generation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:29:36.920808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T14:30:48.687642Z digest=sha256:b442772fa997c3b87d1acc8c1e66fb65258d89e758a97581ae1b35e6799605e2

Observation 73783787-ea88-47f6-8d8f-7347c2ae93c2 · inbound

Towards Error-Free Long Video Generation cites this paper.

Towards Error-Free Long Video Generation Long Context Tuning for Video Generation

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:59:42.128260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T10:49:36.838492Z digest=sha256:4f45c8bf5eaf37c1261dbc3aad74ba57182c9c7d9a3787a17312c484eb77b2f3

Observation 6b89e04f-fe3f-4a2c-9f67-ee55b7245510 · inbound

MemLearner: Learning to Query Context memory for Video World Models cites this paper.

MemLearner: Learning to Query Context memory for Video World Models Long Context Tuning for Video Generation

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.893817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:2f3a59cd5af5ef7afa452970b3ba341bb3484e9842b0999f8aab32dd8f71701b

Observation 9226e234-2594-4550-962c-d8c6a3d48efd · inbound

ICDepth: Taming Video Diffusion Models for Video Depth Estimation via In-Context Conditioning cites this paper.

ICDepth: Taming Video Diffusion Models for Video Depth Estimation via In-Context Conditioning Long Context Tuning for Video Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T17:08:42.920221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T16:59:40.925535Z digest=sha256:cd4500d00fc5f843018c0c5dab2e9bb88ff449770430a59cf022fb719582aeac

Observation 56ea0048-45e1-42c2-9ff5-40f33b1b0d55 · inbound

SlotMem: Character-Addressable Internal Memory for Narrative Long Video Generation cites this paper.

SlotMem: Character-Addressable Internal Memory for Narrative Long Video Generation Long Context Tuning for Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T22:25:25.346355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:25:25.346355Z digest=sha256:8bd307d870b646b6df07aa4e8302c584ac0c3dff69fa795bb54d451be318c95a