Pith. sign in

Paper Citation Record · LEDGER

Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2506.03141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03141 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:59:33.245663Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:58:47.760773Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ce1a6f6f-7ba5-45e6-896e-8926d8ebb6d2 · inbound

Recurrent Autoregressive Diffusion: Global Memory Meets Local Attention cites this paper.

Recurrent Autoregressive Diffusion: Global Memory Meets Local Attention Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T21:59:33.245663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:59:33.245663Z digest=sha256:3af24ab1ee80bdf720017cc48c87c54ec9d5e8d8fdb612fb521dfc300f370630

Observation d1ff0185-3f70-425f-a363-857e497202a6 · inbound

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets? cites this paper.

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets? Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:02:04.221770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T20:00:20.895672Z digest=sha256:d61760fe5f5356a4bf722c31cadebe85cf93a35aeb22746a28ad2889a29e8ebb

Observation dde0d32c-cd0c-4c25-a363-3914c8c1b8c7 · inbound

Vision-Language Memory for Spatial Reasoning cites this paper.

Vision-Language Memory for Spatial Reasoning Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T20:15:36.183372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:15:36.183372Z digest=sha256:af19d3a9958a0b53689b518c48a8585e6f3881a871f53cab422465908826279a

Observation 7d25cc99-4b09-45c2-b852-a67e1d1a6801 · inbound

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling cites this paper.

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:29:56.497402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T14:29:56.348733Z digest=sha256:923b753c9f93f4abd632ebd8c40147ff676208a39f8d64103ba5ca64c54b2e1a

Observation be1124f5-f6ad-4afc-b78c-e2f9f667a430 · inbound

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling cites this paper.

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T16:11:18.301098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:11:18.301098Z digest=sha256:a9fb028e196678b95edfe340c07d0eb3c4f5f944396ca42d3130a96cd272fedb

Observation 0b4530a8-a06c-45de-88d2-bb72ea16e39d · inbound

CustomX: Unified Character, Action, and Scene Customization in Video World Models cites this paper.

CustomX: Unified Character, Action, and Scene Customization in Video World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-03T15:29:00.649927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:29:00.649927Z digest=sha256:dee3313b45e285d6fec99f7737402f4549106f966ee3ca23ab3e70c0b3c0feb0

Observation 6fe4146b-c908-4b1e-93fc-c505b9f41773 · inbound

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models cites this paper.

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T20:36:13.810793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:36:13.810793Z digest=sha256:3ba3fbc1bd0171b1b03fdbeca01940e82366de32d9434ea1de377cd2750dce89

Observation 08ab7c33-aa47-4062-9ef4-dc1d07f78abd · inbound

SuperLocalMemory V3.3: The Living Brain -- Biologically-Inspired Forgetting, Cognitive Quantization, and Multi-Channel Retrieval for Zero-LLM Agent Memory Systems cites this paper.

SuperLocalMemory V3.3: The Living Brain -- Biologically-Inspired Forgetting, Cognitive Quantization, and Multi-Channel Retrieval for Zero-LLM Agent Memory Systems Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:30:50.645013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T19:48:34.606351Z digest=sha256:c63190734ada956558e30445a1b131f51b275cdfd8025ea050621a25dee1c8d7

Observation f841b663-8e70-4ab6-be4e-e21884641ba2 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 281

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.270209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:c6a5df34ef81bc1d870a51e0c2629644699851dbe545359283a1287ebd1eb4e0

Observation 9845b9fb-5bfd-40eb-9c42-9c1f1c7f4afb · inbound

Scal3R: Scalable Test-Time Training for Large-Scale 3D Reconstruction cites this paper.

Scal3R: Scalable Test-Time Training for Large-Scale 3D Reconstruction Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:50:59.685157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T16:56:14.380078Z digest=sha256:56a92fa07129150aa6c13874a1f71aed1364f539f06a499d57cd3c57b7af1f5d

Observation 9bd845dc-2172-4866-a22a-5d35380d1722 · inbound

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds cites this paper.

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:45:27.154100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T13:45:24.961208Z digest=sha256:1084a1c7716d7f6d8e2f9ad025c550dc2e7dbcce46344a7cadb1f28b725f4e01

Observation f4f4daf0-08e0-4e10-9967-4d9d4c0dbb9d · inbound

MultiWorld: Scalable Multi-Agent Multi-View Video World Models cites this paper.

MultiWorld: Scalable Multi-Agent Multi-View Video World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:09:08.126250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:06:11.514186Z digest=sha256:6a7792f82212d0a6c7a1c0f6f33fc04344bb018f053a784423e47a19ba2fd1bb

Observation 036bfba9-9d4c-4994-835c-98b4c678c78f · inbound

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity cites this paper.

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:33:32.128967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T02:31:48.593354Z digest=sha256:21a32481c87ff4c6f4c07ec085d4a036383d879f81e2121b2d4cd70a6547f0dc

Observation 90e471ac-c942-4b50-b2fc-8f0d726cf0c4 · inbound

Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players cites this paper.

Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:53:29.081619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T13:34:03.020051Z digest=sha256:980363652a5f44ff9ec26ae97b4c9c0f3d244b98828f1a27d58b0394fbdf726d

Observation 11b18680-c829-415d-8f2e-fd1bbfbe7f9a · inbound

Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models cites this paper.

Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.130065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T22:43:32.749425Z digest=sha256:0f85364a0f5866d345c75621e487b5a5c0250ccd890c2ebfabe22e32badb459c

Observation 32f981f9-c8bb-43fa-8478-fc258d0b67fd · inbound

DecMem: Towards Minute-Long Consistent World Generation with Decoupled Memory cites this paper.

DecMem: Towards Minute-Long Consistent World Generation with Decoupled Memory Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.789949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T22:38:10.473453Z digest=sha256:c27c0b181eaab49d1548066baf590aa108a64926b295ef00cefdc372dda24c05

Observation 8a09830f-5302-4e12-9481-88fd0609a359 · inbound

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data cites this paper.

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:56:20.240121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T14:52:30.406683Z digest=sha256:2d252ce00ab6efeeb850188808b5656b44c956188e80114ea67ff5f6c25c7e48

Observation 31c4f39c-d9b9-47c6-917c-871c65fc4759 · inbound

Echo-Memory: A Controlled Study of Memory in Action World Models cites this paper.

Echo-Memory: A Controlled Study of Memory in Action World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:57:29.617097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T16:58:37.552036Z digest=sha256:be380f2a68c8906cefd311093ae94f5624685601974a5965fee3095b2af4bef3

Observation 99edf4a0-a104-4d72-b421-d62804aad845 · inbound

Latent Spatial Memory for Video World Models cites this paper.

Latent Spatial Memory for Video World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:07:30.554029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T16:47:42.761342Z digest=sha256:c2d07da6dc1ea17776754665d492a343262cea432561732b044b3b08326128ca

Observation 1086bb37-ba36-442f-93c6-7b29e02e7aba · inbound

FadeMem: Distance-Aware Memory Consolidation for Autoregressive Video Diffusion cites this paper.

FadeMem: Distance-Aware Memory Consolidation for Autoregressive Video Diffusion Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:37:37.482634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T13:44:44.550033Z digest=sha256:c5bfa8a4556dbccac0f7eea23571c962819a1cb96e3fafc00ef682b9242af952

Observation 6d2f390b-73ae-473a-8785-784f244508a6 · inbound

PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory cites this paper.

PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:58:47.762544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T03:23:58.455778Z digest=sha256:794452d668d684e787707f26cfc579f46ba0ad96d17b45dfce4b725f7f1836be

Observation 7bf165a4-308a-412f-8e7b-143ac53ad726 · inbound

MemLearner: Learning to Query Context memory for Video World Models cites this paper.

MemLearner: Learning to Query Context memory for Video World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.877273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:3fef428ef593dfdbb8c321373ff42d82fda145c85b76f8ab40f10148ec059f3c

Observation 8e84fb0f-18ac-4197-8c51-0b5170118c76 · inbound

WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory cites this paper.

WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:31.047823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T14:28:24.922144Z digest=sha256:1f72d412a9e81d2010117b464965c7705b2204fd0e2b5c55439b2352c10eece8

Observation d03d8117-f012-4c67-ad39-cdec4511a596 · inbound

SlotMem: Character-Addressable Internal Memory for Narrative Long Video Generation cites this paper.

SlotMem: Character-Addressable Internal Memory for Narrative Long Video Generation Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T22:25:26.677846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:25:26.677846Z digest=sha256:f7aa2cbc535a22bec786e64cb1783755e60defa15f4331282b052b0b15f1ddc4

Observation 8b85fad5-f630-4ea3-bbe1-05ccf3774b37 · inbound

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation cites this paper.

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T10:37:56.086456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:37:56.086456Z digest=sha256:80cf0a39c39f0c52e37b847682eccba8541270ce88ae4731b1067b4bf498caa8

Observation b9f5f915-f899-4ff6-bc6a-ddad2266ea8e · inbound

Wonder: Video World Model Done Better cites this paper.

Wonder: Video World Model Done Better Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T00:52:27.986020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:52:27.986020Z digest=sha256:83074b44563be2e80d87e6e48d6484684d51eb55f5b9720289807d038b307f4c