Pith. sign in

Paper Citation Record · LEDGER

MemLearner: Learning to Query Context memory for Video World Models

As of 6 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 1 inbound Pith citation observation for arXiv:2606.31734.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.31734 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-01T05:30:56.140465Z

measured 76 of 76 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T06:35:29.083742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

75 of 75 outbound references displayed

  • verified exact14
  • verified fuzzy19
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch32

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d8c796a4-132c-4d00-9fa2-4fd248f67f1f · outbound

This paper cites MAGI-1: Autoregressive Video Generation at Scale.

MemLearner: Learning to Query Context memory for Video World Models MAGI-1: Autoregressive Video Generation at Scale

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:25:41.913503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:f6bb6ebc29b5b3ec647dc89617ce68e80d43a80e9f6eb5cc6514e6e0d05ed137

Observation a7317523-9916-46a6-a70a-de4ea7b69bee · outbound

This paper cites Advances in neural information processing systems35, 23716– 23736 (2022).

MemLearner: Learning to Query Context memory for Video World Models Advances in neural information processing systems35, 23716– 23736 (2022)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.157923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:29aac68c0212f18125645559888b1b339ff6673d9c11ef0f894e72f89b054fa0

Observation 18343c31-284e-4b4d-bcea-e216b7bff1d3 · outbound

This paper cites Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames.

MemLearner: Learning to Query Context memory for Video World Models Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.855643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:b703fcc36319691000c202ca7551c37fc35125764128c14a300ce7eb00f37716

Observation c4a367f4-f10c-4e9d-8bd4-12721e2f95bb · outbound

This paper cites V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.

MemLearner: Learning to Query Context memory for Video World Models V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.906075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:5cd8b3d508cd6806bdb40c7d44039bc8701eca5b9a392445f5d8f650c5c76e22

Observation 10e80f2b-a20f-46ef-b919-bdd34158fde6 · outbound

This paper cites ReCamMaster: Camera-Controlled Generative Rendering from A Single Video.

MemLearner: Learning to Query Context memory for Video World Models ReCamMaster: Camera-Controlled Generative Rendering from A Single Video

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.908158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:0eaa4a8e518d951d5a494fb56a15b253798316dcd1c971f151959e1d03cf553b

Observation f5639ada-9b07-40f7-9a3d-e245feec0764 · outbound

This paper cites Yu et al.

MemLearner: Learning to Query Context memory for Video World Models Yu et al

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.150219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:97608f271b399679c2ac823e167ff9d5905f3fe664740abea94cf4cf12bea2e5

Observation 23814f50-bb84-46bd-bfab-beb691e994f2 · outbound

This paper cites Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models.

MemLearner: Learning to Query Context memory for Video World Models Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.915950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:5f63f5f9d7d19fad95fcb1f93f77ae24e247385323f6f7f1b28a931e1a65768a

Observation 78791b75-3313-4bd5-ab33-f70ab5e57118 · outbound

This paper cites Advances in neural information processing systems33, 1877–1901 (2020).

MemLearner: Learning to Query Context memory for Video World Models Advances in neural information processing systems33, 1877–1901 (2020)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.168179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:30bde9994d5b0308de8c2c0e02fa5c9410422277c50e0e7ce41dfb8774741500

Observation ac3e153b-5d67-49c9-952a-95897ad4d431 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.175612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:f5c9e0f660559c514b017f09bbded87a1f7fdbf75b729a36716ce3a9b148255e

Observation da3f998c-412c-47bc-8cd9-61226e1631d6 · outbound

This paper cites In: arXiv (2025).

MemLearner: Learning to Query Context memory for Video World Models In: arXiv (2025)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.170029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:2f4c91e9e592fb4a15d97e4fed35d02530d2c4bc91f723e614a9c39d69ee791f

Observation 9a926ef8-df17-4961-bc3d-caf26cc92a50 · outbound

This paper cites Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion.

MemLearner: Learning to Query Context memory for Video World Models Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.905323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:0d226007dc8f238631a4c93accd9cc72ad4b79044a95319b3aa450f5da5927ce

Observation 75c8dab1-c976-4b09-922f-1389cc1c5fa0 · outbound

This paper cites VRAG: Learning World Models for Interactive Video Generation.

MemLearner: Learning to Query Context memory for Video World Models VRAG: Learning World Models for Interactive Video Generation

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.910723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:13de39487c83e4b47126e034ff21b659e7af1807552ec34620c54c15497e5409

Observation 2e29b78c-b609-4728-83b4-357d09374928 · outbound

This paper cites Self-Forcing++: Towards Minute-Scale High-Quality Video Generation.

MemLearner: Learning to Query Context memory for Video World Models Self-Forcing++: Towards Minute-Scale High-Quality Video Generation

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.903498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:e25aaf76c819ec761e3631637192d159eda6e5c1e0f971c0f240e7a58e34e483

Observation c98223a1-3a5a-450e-be53-536016f92f0f · outbound

This paper cites Emu3.5: Native Multimodal Models are World Learners.

MemLearner: Learning to Query Context memory for Video World Models Emu3.5: Native Multimodal Models are World Learners

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:25:41.921325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:d665786801be96da21e45e327c3cbf8d831a11e08be3f557958ea4108b79c1dc

Observation 9fb31033-1492-4bf4-bd9f-847bddc84627 · outbound

This paper cites In: Proceedings of the European Conference on Computer Vision (ECCV) (2018).

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the European Conference on Computer Vision (ECCV) (2018)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.163093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:e79cd7e5319b2fec0aee29c20f90cafdd8e6ba3390421f6159b40f9996a1d732

Observation b199cec6-ad40-408a-9f81-9d756eafff11 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.165957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:fd7e294a052d5b136ce3c1136af6c6a2a734e6b55752912cfb89bebd05647f17

Observation 9d0de4d9-f4e4-46ca-9ee1-936be45d5336 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:56.378054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:fc2b03c9ff67ae1263d742a5be3015ceb9230cb654d2bffb514eea5b7d02f860

Observation c999d9ed-8cc8-44a0-9c0c-b14696254935 · outbound

This paper cites Autoregressive Video Generation without Vector Quantization.

MemLearner: Learning to Query Context memory for Video World Models Autoregressive Video Generation without Vector Quantization

Reference 18

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.898990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:54def3aa4e6888905ac0cb2d2d46a2f10a0197bdbf22f860a05938edece53603

Observation 84cba3e3-b30b-4ead-9651-679b279b5dae · outbound

This paper cites In: ICLR (2025).

MemLearner: Learning to Query Context memory for Video World Models In: ICLR (2025)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:42:55.781411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:dafec55bad1cc2d7e2784cffa38b585ab6a4a1f8c3ee0895053689d4cc160e7f

Observation 54ef385e-cd60-4df5-b5f5-ebede4a10af2 · outbound

This paper cites Long-Context Autoregressive Video Modeling with Next-Frame Prediction.

MemLearner: Learning to Query Context memory for Video World Models Long-Context Autoregressive Video Modeling with Next-Frame Prediction

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.883012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:a9b3fb30ceb9070b885777d36a5e03bb74eeb2f16fdb0d1c765050e60bcf965c

Observation 6b89e04f-fe3f-4a2c-9f67-ee55b7245510 · outbound

This paper cites Long Context Tuning for Video Generation.

MemLearner: Learning to Query Context memory for Video World Models Long Context Tuning for Video Generation

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.893817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:2f3a59cd5af5ef7afa452970b3ba341bb3484e9842b0999f8aab32dd8f71701b

Observation 82bb8e3d-c4e3-4070-b98c-b0dec0973ee0 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:56.363212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:772aa3ac34af9563e156a03e9c4f3be968397f78b09ef963a6c3084b3a1d1543

Observation aae2484c-f770-4b70-a50d-812a7bb414ca · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

MemLearner: Learning to Query Context memory for Video World Models CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:25:41.908460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:55bd591608438be69f151381c627867ae892c86699fb61014bb5b3c0f89e6e0b

Observation afa6254d-dba9-4036-803b-59356d35d420 · outbound

This paper cites Classifier-Free Diffusion Guidance.

MemLearner: Learning to Query Context memory for Video World Models Classifier-Free Diffusion Guidance

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.892002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:6506ee2ae125b7fd2e29db511e436f16f28dc042da423be7108c207455ed65f8

Observation cae09e41-14cb-401d-8834-fe323949b14b · outbound

This paper cites Relic: Interactive video world model with long-horizon memory.arXiv preprint arXiv:2512.04040.

MemLearner: Learning to Query Context memory for Video World Models Relic: Interactive video world model with long-horizon memory.arXiv preprint arXiv:2512.04040

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.897139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:37302dd294c0194256a5650a7470725da4052c7eb05e179d7be494edef8b1593

Observation 12793a0a-e0bb-41bd-8971-dd9f8a2f5ca2 · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

MemLearner: Learning to Query Context memory for Video World Models Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.899992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:631421eef79b6f571d294707111f8ed141213aa73f3746d0c6d2d852c172a11a

Observation a18952b9-13f4-40bd-bd17-3d811a25faa6 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024).

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.153991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:fc819c341f8f335fc8bc13d2699bcf7983b863ce56e08aa9e6a9235aafc1316f

Observation 8a4e114c-3eef-4c31-900f-bafe3caed023 · outbound

This paper cites Nature638(8051), 656–663 (2025).

MemLearner: Learning to Query Context memory for Video World Models Nature638(8051), 656–663 (2025)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.152130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:7e46e2dcf9ab67fd32f508f052e2b0971bb10f2ab359e24281449841a6068e24

Observation 681661fa-5577-4f48-b633-dbc5863dae0c · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.144449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:a58d866ac41f1ed4c020ec0f2422d598d9c2b9a9cec9f2477f48d9df5b929716

Observation 18414c1c-451c-4e81-bd55-b39d542d5a54 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.146414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:2623b09624f9f5ed8863afcc2f8a37a6013bb29136d3c6d818ce18db94ae76ba

Observation b9ec92f2-ce38-4bc6-bc7d-c4d9638dab89 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.148291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:aff80b9a7e4555908df98dfec580e48eb8be839876c5d3b47d6a4dfef158181b

Observation d09c0f50-9b73-4199-9535-72598503a29a · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

MemLearner: Learning to Query Context memory for Video World Models VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.910873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:b6d0ded358b4ea84e682cb634772756f804aa314a3f0d3ff148f9f87f6ef4d45

Observation 5194b4d8-d4c6-4d11-ad92-971724e51a0b · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

MemLearner: Learning to Query Context memory for Video World Models HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.913335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:c71a3e22d82eaa9e620d2255228cc95467a1e4d068eaed96be50575fd3326752

Observation 8c5e2fdc-c89a-451b-aa6f-839914c1cc71 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.155823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:eb529a5cc7deb191a9412f49fe6151eea455c3d52064a1e6a1d22ad8e67d2ab1

Observation 816534a7-565c-4a14-9596-9814748547c7 · outbound

This paper cites In: International conference on machine learning.

MemLearner: Learning to Query Context memory for Video World Models In: International conference on machine learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.159902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:16f5d3a2d58f97da54f612b3310948f1db9e5112e66745593cc3fc1417d4d19d

Observation f7f4d308-7a1d-48e0-b60e-b026f81abe37 · outbound

This paper cites VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory.

MemLearner: Learning to Query Context memory for Video World Models VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.888623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:22a00ba523fcbb908fd2a1a597b38f5c9953d76cee69b144738a091ecf9d059c

Observation 1ac16127-092b-4769-94aa-2905eb528ac3 · outbound

This paper cites Autoregressive Image Generation without Vector Quantization.

MemLearner: Learning to Query Context memory for Video World Models Autoregressive Image Generation without Vector Quantization

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.896692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:ddbb3b79c2b7e6f483332410e2cee2da02f074feffd7e69c9c56755fc3ea7ac9

Observation 17dbca03-176d-42f7-8cd7-2043e9a48df9 · outbound

This paper cites Stable video infinity: Infinite-length video generation with error recycling.arXiv preprint arXiv:2510.09212.

MemLearner: Learning to Query Context memory for Video World Models Stable video infinity: Infinite-length video generation with error recycling.arXiv preprint arXiv:2510.09212

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.861034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:82c596e67497e5a0a98c05d7087218c5b84f074a098d0941cf11cba2e8d69c5c

Observation 7e34d0c9-71cf-440b-b908-fdd8adc2fdf6 · outbound

This paper cites Sekai: A video dataset towards world exploration.

MemLearner: Learning to Query Context memory for Video World Models Sekai: A video dataset towards world exploration

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.849473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:d77172223414451d16953a37970c3c569fd712c3d6214051cf5f1fce6ad3752c

Observation 67c37e6f-b2e7-49d8-83d2-9be801eb0453 · outbound

This paper cites In: Proceedings of the IEEE/CVF international conference on computer vision (2025).

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the IEEE/CVF international conference on computer vision (2025)

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.140613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:5ad90dcb9ff905d520d902a6d9d67e98e1dc9cf60158770ffc2e422bb64a5cec

Observation caa55e36-23fa-416d-8b9d-c8db606074da · outbound

This paper cites Rolling Forcing: Autoregressive Long Video Diffusion in Real Time.

MemLearner: Learning to Query Context memory for Video World Models Rolling Forcing: Autoregressive Long Video Diffusion in Real Time

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.915944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:50b7e998327ec189a3973e67c8d1e3697d88f1afd53edf6ec073f71535c201ff

Observation 7fb6c6e3-cebc-402a-b540-9085c512c474 · outbound

This paper cites You See it, You Got it: Learning 3D Creation on Pose-Free Videos at Scale.

MemLearner: Learning to Query Context memory for Video World Models You See it, You Got it: Learning 3D Creation on Pose-Free Videos at Scale

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.860876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:ca9a3dc0a0046db64f168c0852e149d172b84d5c448a76bc7e1386b531e5357c

Observation b2c96aa3-b62a-4177-8817-3d0a47eef048 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.142488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:48fa2799d8ac75c9728ac7c11199a33223be285e35f0296b06dde3bea79d93dc

Observation 81a84eb7-8fb1-4ae8-99f9-ead925a67836 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision (2023).

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the IEEE/CVF International Conference on Computer Vision (2023)

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.132982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:72f7637ebf8edbb2445ad8c1c25c215990d8bc88a86dd56fcbcc43f46bb5c30a

Observation aa774030-1094-409a-b8ea-c2344f10c2e0 · outbound

This paper cites Long-Context State-Space Video World Models.

MemLearner: Learning to Query Context memory for Video World Models Long-Context State-Space Video World Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.872252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:27401b8abe219ac34fa334d32de05f079c4a7976f3d819f02696c693b21be556

Observation aea3eaaa-479d-42eb-9dd9-6574935f633c · outbound

This paper cites In: International conference on machine learning (2021).

MemLearner: Learning to Query Context memory for Video World Models In: International conference on machine learning (2021)

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.134912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:b761c042a90dc7bc346e84dc7ea5919b37a8f4646d7121b8c22d80bebfbb0186

Observation 04f205f9-923d-4577-9eff-6ed2399902e5 · outbound

This paper cites GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control.

MemLearner: Learning to Query Context memory for Video World Models GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.880404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:0427c4745d4232b38ebe259dfb3aaf7ffd1316060d2565f6187356486fa3be50

Observation b987a8e7-7a66-4969-a198-544585e4ffb1 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.138753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:57d44e20b8a97ae2fdde34595a6f6217fae485e57056b1f9e84ceaff6fd4f4f7

Observation 84c63bd2-aaf7-4a3c-b20f-6dfe87c726f5 · outbound

This paper cites History-Guided Video Diffusion.

MemLearner: Learning to Query Context memory for Video World Models History-Guided Video Diffusion

Reference 49

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.919041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:1385d292c5e6bccb2505d1fde905a1a0ebfa52e11b82c3c3d8270c3a66fc4d73

Observation 6d647461-2589-4cc2-a54b-a8e47a753a79 · outbound

This paper cites WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling.

MemLearner: Learning to Query Context memory for Video World Models WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling

Reference 50

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.853105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:3acde9b013cb88cfba4b6a277e0060fb0f4b1a49f83d1f545fec47bec1382ea3

Observation 0432838b-d6ee-487d-9491-890efb8ee03c · outbound

This paper cites Diffusion Models Are Real-Time Game Engines.

MemLearner: Learning to Query Context memory for Video World Models Diffusion Models Are Real-Time Game Engines

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.891328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:f6b8097800e304af0ea1266de278954c1b903e5d34c01a5a50912b1edde3f7d6

Observation 12624228-3b4e-4d3e-a2d3-55da533fbc2c · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

MemLearner: Learning to Query Context memory for Video World Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 52

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.836084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:6519bc06a97ca4bc68a2423b4d42f80c34bb48a46c05d67b5f1b746524ad9342

Observation d07a1447-45c7-45f5-9389-ca870f8e7605 · outbound

This paper cites Spatialvid: A large-scale video dataset with spatial annotations.arXiv preprint arXiv:2509.09676, 2025a.

MemLearner: Learning to Query Context memory for Video World Models Spatialvid: A large-scale video dataset with spatial annotations.arXiv preprint arXiv:2509.09676, 2025a

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.830907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:b5c6c7de6d61e9d58ca00378fc10f16a49095d322e2fc4f5aca06c579c6040a4

Observation c33ce58a-702d-4ced-b31d-5e2776be1de2 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

MemLearner: Learning to Query Context memory for Video World Models Emu3: Next-Token Prediction is All You Need

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.833228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:82a46b255136fd4583c0488e17213823f88df0dda0e4bfb9f7a122ee34e1b7ff

Observation 2bba5cca-22dd-49eb-b132-e98335b05796 · outbound

This paper cites In: ACM SIGGRAPH 2024 Conference Papers (2024).

MemLearner: Learning to Query Context memory for Video World Models In: ACM SIGGRAPH 2024 Conference Papers (2024)

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.126584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:b61686848871133f9be179d91451f1fc3e416a19915d59f50b6b86fc7609d32c

Observation e4e4f89e-2fa9-4b30-8169-d9633376bce8 · outbound

This paper cites Video World Models with Long-term Spatial Memory.

MemLearner: Learning to Query Context memory for Video World Models Video World Models with Long-term Spatial Memory

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.838697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:d9d72bbb81e51cdf2dee77252f5db2d62a0bffc09617bf854e821ab82e234863

Observation 2709d5c1-48e3-48ab-a179-a185295505ec · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.128817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:b8919458253fec7c8c73c78e724205a0f11702fd47c7357a3c6a1933043ef6b4

Observation 265ec9f6-814f-4f2b-8463-8bdcce49c448 · outbound

This paper cites arXiv preprint arXiv:2504.12369 , year=.

MemLearner: Learning to Query Context memory for Video World Models arXiv preprint arXiv:2504.12369 , year=

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.846449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:ca5a21cbfdbab227c98145b1fab24aee683b73d4d344faf338b9dacd8dffcf98

Observation f33a7c57-0642-4630-bc56-ed301e2b561d · outbound

This paper cites VideoGPT: Video Generation using VQ-VAE and Transformers.

MemLearner: Learning to Query Context memory for Video World Models VideoGPT: Video Generation using VQ-VAE and Transformers

Reference 59

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.872227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:1a7e17b435f58e22505ffc7a09d1ba81a7b778eea5d49cc08be478a9ef347572

Observation afaa112b-8b63-4005-b62b-ca90d465f046 · outbound

This paper cites In: Proceedings of the 41st International Conference on Machine Learning (2024).

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the 41st International Conference on Machine Learning (2024)

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.122371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:5d4d18e153c365da227a8ef104bec26b2128ef529cecf2544a85401f91f49d2b

Observation 8746157e-ccd6-4889-bb38-21511a3953ef · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.120347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:cad326837db5c1c12b34cbf6cdd774636f0bc3d8ad97df8453f59e61a9b8dfbf

Observation a9bb7f45-bd7e-40c9-bcc9-e66f8a66cfa0 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

MemLearner: Learning to Query Context memory for Video World Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.828161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:5b34dd4b94ddbb0db63b81e74ed3eae51e18d5ee353def1caabe33737fc60344

Observation 17abb7ed-3b89-4867-be70-77e2668940d6 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

MemLearner: Learning to Query Context memory for Video World Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 63

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.817110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:cd5e0bf83409f3c10ae47b9813f68e3c4836954c16354c2222ae75ca9c15e898

Observation 393f630b-cfc2-4b5a-a16f-a387359afdab · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.124546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:fb8ed690f066b9a78a64574f9afa513844be05d130e03dc89e5c2bb964a5bc26

Observation 44e5c282-180f-4269-ae33-5536e43c23dc · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

MemLearner: Learning to Query Context memory for Video World Models In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T21:32:57.130980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:2b2f1ddbe59c52a4cc6c97c1d2d491998d008969801222541e424793d95fec3b

Observation 7bf165a4-308a-412f-8e7b-143ac53ad726 · outbound

This paper cites Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval.

MemLearner: Learning to Query Context memory for Video World Models Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.877273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:3fef428ef593dfdbb8c321373ff42d82fda145c85b76f8ab40f10148ec059f3c

Observation 764f8600-24b9-4a67-b451-03bfe4c7113a · outbound

This paper cites A Survey of Interactive Generative Video.

MemLearner: Learning to Query Context memory for Video World Models A Survey of Interactive Generative Video

Reference 67

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.811839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:da82e2824c76b5f6efca870a8decbc809e3db7e92930815f1166f1abf13f7402

Observation d87793ea-c4cc-475f-a2fa-983dfdeb384c · outbound

This paper cites Position: Interactive Generative Video as Next-Generation Game Engine.

MemLearner: Learning to Query Context memory for Video World Models Position: Interactive Generative Video as Next-Generation Game Engine

Reference 68

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.819798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:2b4d1431517c14a1aeb23d6b6e756444410207ce1c69c80191e5b1550334e3c0

Observation c8295003-00f8-4813-a388-0e2464ae6b08 · outbound

This paper cites an unresolved cited work.

MemLearner: Learning to Query Context memory for Video World Models Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-07-06T21:32:57.136862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:2d9b1367e96488ad92634495fb8f0770a7b60b8aee75bb9b26cd0304bbf28a7d

Observation 1b1944bf-bab8-4e10-aa3c-0cda17ff2167 · outbound

This paper cites ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis.

MemLearner: Learning to Query Context memory for Video World Models ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis

Reference 70

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.810794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:040a64c09a609f41e5c6c11814330184395c9ef0c6e958f0d00f030d923327dc

Observation 9c4e45ad-e197-411c-a260-d00cdf4cb04b · outbound

This paper cites Lvmin Zhang, Shengqu Cai, Muyang Li, Chong Zeng, Beijia Lu, Anyi Rao, Song Han, Gordon Wetzstein, and Maneesh Agrawala.

MemLearner: Learning to Query Context memory for Video World Models Lvmin Zhang, Shengqu Cai, Muyang Li, Chong Zeng, Beijia Lu, Anyi Rao, Song Han, Gordon Wetzstein, and Maneesh Agrawala

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.874804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:ccea294c5212231c926e69a198a523bd48ce8ae9fa3ec5655e20b15a7dee258e

Observation 3e632ca4-59da-4162-9a96-2926a0c6a59e · outbound

This paper cites DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning.

MemLearner: Learning to Query Context memory for Video World Models DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning

Reference 72

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:25:41.850838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:ddc890cf6c884e4fa858269179df542a6263ea4cf769673de245fcdc14876fb2

Observation fd39595d-301a-4a34-b854-dfaa4c1b5689 · outbound

This paper cites Learning 3D Persistent Embodied World Models.

MemLearner: Learning to Query Context memory for Video World Models Learning 3D Persistent Embodied World Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.879808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:f5808dd0122b0312ceb0287fc9abb982cfa6cee59f000a70ec92eeef927c8d43

Observation ea32659f-6a06-4cf6-8112-b07271200be9 · outbound

This paper cites Omniworld: A multi-domain and multi-modal dataset for 4d world modeling.

MemLearner: Learning to Query Context memory for Video World Models Omniworld: A multi-domain and multi-modal dataset for 4d world modeling

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.837376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:5491cb618b195bc6b35a54b863973e35bb977d78eb7b929a58b46099c1945e0f

Observation cae434dd-a0cc-4eb1-9612-8db8fb619020 · outbound

This paper cites IRASim: A Fine-Grained World Model for Robot Manipulation.

MemLearner: Learning to Query Context memory for Video World Models IRASim: A Fine-Grained World Model for Robot Manipulation

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.866546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:8f84ecee2b741ed05a78d2a4f0673649831576b6f253f6705c16ffe3c9253566

Pith citing papers

Observation 32357189-8840-4fc4-92d5-5f079e1f3ae0 · inbound

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering cites this paper.

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering MemLearner: Learning to Query Context memory for Video World Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T06:35:29.083742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:35:29.083742Z digest=sha256:a9af99eddd3d7ace7b06440c6207fdf2a0a463edcffde55f1f12d7ed8e49523d