Pith. sign in

Paper Citation Record · LEDGER

WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2401.09985.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.09985 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:31.307028Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T07:26:54.481828Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a403d5d3-3523-41dc-97ee-a0e3898b796d · inbound

ProphetDWM: A Driving World Model for Rolling Out Future Actions and Videos cites this paper.

ProphetDWM: A Driving World Model for Rolling Out Future Actions and Videos WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:31.307028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:31.307028Z digest=sha256:4e90a0217caa29719d7f63d07c7827afc3ba8d2fab67f72c7c0d754f69aba657

Observation 909be15d-779c-4c8f-9a89-a3e915be256c · inbound

Long-Context State-Space Video World Models cites this paper.

Long-Context State-Space Video World Models WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:03:20.562826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:03:20.562826Z digest=sha256:10542466ac23dd57b20788656ca7965f8e8e31f3631b10540a9b4947a62f0591

Observation 070a908b-344a-46e1-9682-121b51bcfe98 · inbound

GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control cites this paper.

GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T13:13:46.886225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:13:46.886225Z digest=sha256:fb5bf1f92d3d57f7009169024358cb58b4360e80e4126560ee48186251823213

Observation 5527bad4-bd6c-4d2a-a004-2ab86e607bdb · inbound

MOVi: Training-free Text-conditioned Multi-Object Video Generation cites this paper.

MOVi: Training-free Text-conditioned Multi-Object Video Generation WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:00.850247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:01:00.850247Z digest=sha256:f33103ea27c9a12320af1d1e5739e8df5ad1a9de62c15d75ea243abfa69fa6a6

Observation c2ba2f92-7369-45f0-b211-78787180e80e · inbound

World Models for Cognitive Agents: Transforming Edge Intelligence in Future Networks cites this paper.

World Models for Cognitive Agents: Transforming Edge Intelligence in Future Networks WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:36.336903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:08:36.336903Z digest=sha256:aefc5c9cd713ad29bf0bc66c2bec23ff2030401a4e8e10b71b16d03cf0837320

Observation 02144c37-fea5-499c-b8e0-7e0622570869 · inbound

Hunyuan-GameCraft: High-dynamic Interactive Game Video Generation with Hybrid History Condition cites this paper.

Hunyuan-GameCraft: High-dynamic Interactive Game Video Generation with Hybrid History Condition WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:04.283988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:04.283988Z digest=sha256:3c6fc5e17389072fee1947cc90dbb5c4ccc5bf6f29103a81296fa32223b06ef5

Observation d07a90e7-d50c-4e12-aa4a-6cac09631bd7 · inbound

EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling cites this paper.

EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T19:35:42.847397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:35:42.847397Z digest=sha256:e0d286108842918ef58e37181b074d4918a09f332f5b47534cd2367bb7c1c7aa

Observation 5a0906ad-ae6e-45c4-98de-e213bc1142c3 · inbound

World Model-Based End-to-End Scene Generation for Accident Anticipation in Autonomous Driving cites this paper.

World Model-Based End-to-End Scene Generation for Accident Anticipation in Autonomous Driving WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:06.168043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:06.168043Z digest=sha256:8092bf2a4bf49a8baadc9e25e966f656a7491db3d18393a1ddf74ceeab67c921

Observation fe0eb035-5071-4400-b47a-e6a17f6bbff8 · inbound

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models cites this paper.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:02.986657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:02.986657Z digest=sha256:bc173fe6dd577367f4446c73e2d35407cb5726de3e189df7507456e0089e4bf3

Observation 4811cc60-00c2-4d3a-855c-db6f784593b1 · inbound

Bounding Distributional Shifts in World Modeling through Novelty Detection cites this paper.

Bounding Distributional Shifts in World Modeling through Novelty Detection WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T22:58:41.359450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:58:41.359450Z digest=sha256:001df7580a730d864e29ee4a84f147d32650546ffb17c13979e96cbadcb6b3af

Observation 2cb8cdbc-7c98-453f-8dc6-0ae7e0e337cc · inbound

Ego-centric Predictive Model Conditioned on Hand Trajectories cites this paper.

Ego-centric Predictive Model Conditioned on Hand Trajectories WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T15:29:25.564545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:29:25.564545Z digest=sha256:434ce8962e26f3eeb930b7bb984af2b0b6c7c55337393ec329023982ceab7089

Observation 2379bea7-79bb-4acc-860c-822a9dae3eeb · inbound

A Comprehensive Survey on World Models for Embodied AI cites this paper.

A Comprehensive Survey on World Models for Embodied AI WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:34.159128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:12:34.159128Z digest=sha256:9042ff989991b732e9683f5f61df14e2b406922e77bca3fac634106e4134c93a

Observation 26ef742e-c6f0-44ef-a0e4-6df9e6d97bf6 · inbound

RynnVLA-002: A Unified Vision-Language-Action and World Model cites this paper.

RynnVLA-002: A Unified Vision-Language-Action and World Model WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T20:59:52.965693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:59:52.965693Z digest=sha256:39bdd1b462c4e3bdfe3d3cd41aaa95e5828cf9d9de7ee34771c15cddcf73c83a

Observation 651d5fd1-9608-495e-86c4-a9fdda929ac9 · inbound

Learning Vision-Language-Action World Models for Autonomous Driving cites this paper.

Learning Vision-Language-Action World Models for Autonomous Driving WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:31:00.630911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:08:10.442655Z digest=sha256:4e71f267934e5157d079f68738ddc17723116932f462d300e3601910a8dc442b

Observation fb69bb43-94cd-4a36-832a-1e1c72498613 · inbound

MultiWorld: Scalable Multi-Agent Multi-View Video World Models cites this paper.

MultiWorld: Scalable Multi-Agent Multi-View Video World Models WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:09:08.137060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T05:06:11.514186Z digest=sha256:729143fd4f11219e63e8e45286b3f97f9476161e52dbb4431fa41193ca788a04

Observation 1b1da9b8-f329-4cee-aceb-c00d0689b945 · inbound

TRAP: Tail-aware Ranking Attack for World-Model Planning cites this paper.

TRAP: Tail-aware Ranking Attack for World-Model Planning WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:04.888141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:15:45.407680Z digest=sha256:05d67348880b5c3d34001d431cf01499a6b99fe628aba64030dc0c7a50b5e684

Observation e971dfd0-1021-4e5d-9cf2-b9d926340efe · inbound

From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation cites this paper.

From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:57:17.900778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T04:44:03.661688Z digest=sha256:46bf47626cce69f757171831935ced1d5a575526d2a84e7a788f774c5ad7f2a1

Observation 053eab7f-6169-4bd4-8288-7e3dfe9b8472 · inbound

GeoWorld-VLM: Geometry from World Models for Vision-Language Models cites this paper.

GeoWorld-VLM: Geometry from World Models for Vision-Language Models WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-20T17:58:49.588513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T17:57:00.909897Z digest=sha256:965c929bf1ff149903438b6166735b505911b2d623924f103e787da7ba8dbbe0

Observation 154bcbdb-bb8a-4954-bc18-176528501ffe · inbound

GeoWorld-VLM: Geometry from World Models for Vision-Language Models cites this paper.

GeoWorld-VLM: Geometry from World Models for Vision-Language Models WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:05:00.616391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T19:02:05.125937Z digest=sha256:e96aae9deff5d4291b8bb75a5709b6e6026c57a63d80ffd3ac39f25499ea57b4

Observation fcc281d8-3e7e-4465-8d2d-8d5739501636 · inbound

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment cites this paper.

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:16:44.709304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T07:02:37.291472Z digest=sha256:c7b6a1641d90b699c2a8de8955760b1e6fb55853cc80a96e17bc2dd4445c7af5

Observation 6ecb5a2e-52ad-48f6-8d59-35cac1e89ea9 · inbound

GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation cites this paper.

GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 121

Resolution
unresolved
no resolver link, observed 2026-07-12T08:04:48.963890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:04:48.963890Z digest=sha256:246a0891caedc9672083a74098746a91491a77da74f2a726e14e4c7ccb7a4ae1

Observation a0789bfe-3947-42dd-82ac-bd6ef23cb264 · inbound

Targeted Structure Completion for Sparse-View 3D Reconstruction in Autonomous Driving cites this paper.

Targeted Structure Completion for Sparse-View 3D Reconstruction in Autonomous Driving WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-11T15:37:01.397370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:37:01.397370Z digest=sha256:7efdfe654579312085bfb307aeaf173a9e26fb4f028916b4a6829f9896371152

Observation 92a24d95-39fc-4534-b703-ebd4576db2bf · inbound

Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents cites this paper.

Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-10T07:26:54.483224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T07:18:57.823444Z digest=sha256:0e2ca13ecae5675e164b845e7b090d5e9a82cb79741a187ac2208c0081fcd9d3

Observation 597caa6f-e5d4-4498-8e63-5e4e3ee1775b · inbound

Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents cites this paper.

Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T07:56:56.491230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:56:56.491230Z digest=sha256:4a8fb1df02ac731faca5de291e052e6cf7ecb1b18df391fcb6a38281f68ca0be

Observation 0f539b69-ceb2-4482-bd59-174de125e61f · inbound

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch cites this paper.

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T03:16:53.075596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:16:53.075596Z digest=sha256:7d51170a1271256041a9b9f7230f73526e2d09d8908cccfce53e0a76b49a758f

Observation 5bbb7262-ef3d-492a-be20-e2d399360cd3 · inbound

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories cites this paper.

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T00:04:08.689523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:04:08.689523Z digest=sha256:d6600d2cd6b8dc72efdd2689560556938ce2492fd188d58d2824d57e057ccb46

Observation a8405e07-acc2-4b16-bcb5-82b027e61bed · inbound

DeforM: Reasoning-Guided Physics-Aware Video Generation via Spatial-Temporal Masking cites this paper.

DeforM: Reasoning-Guided Physics-Aware Video Generation via Spatial-Temporal Masking WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T14:44:20.192036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:44:20.192036Z digest=sha256:c5bce0f3ab1eb41dabe7eb5fed5cdb6683621b0fbba9ad69e686bb2e3f3f5e04

Observation 08b1164e-9f3f-4ef8-a680-bade7b797e1f · inbound

LeapBot-WA: World-Anchor Action Models via Predictive Latent Alignments cites this paper.

LeapBot-WA: World-Anchor Action Models via Predictive Latent Alignments WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-31T23:29:12.915500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:29:12.915500Z digest=sha256:40c13245bb772e8031daa8ee243cbba9258078f0af8a1caedd8c5fd301d91752

Observation cda22c50-5f3f-4e2e-91de-a2363be01409 · inbound

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation cites this paper.

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-30T21:05:52.754117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T21:05:52.754117Z digest=sha256:04ac953e081d453887326195487bc190959bb6b30bab55c2ca8904c7eeedf0a4