Pith. sign in

Paper Citation Record · LEDGER

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models

As of 12 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2608.06799.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06799 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:41:43.912335Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact2
  • verified fuzzy2
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c0ab036a-f408-45a3-84df-45276b015293 · outbound

This paper cites Ctrl-World: A Controllable Generative World Model for Robot Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Ctrl-World: A Controllable Generative World Model for Robot Manipulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.805937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.805937Z digest=sha256:be13a07f5f96334af91b87b0414e96309ac8f159a3905955e423ca0b4fc7675c

Observation ff021cfc-2268-44d6-ab40-4170ef71fb38 · outbound

This paper cites World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models World Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.810581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.810581Z digest=sha256:249cf6f841e564089ca30551bbb1ad909c5aedd00a75cc0981fdbcbcb1f73cad

Observation 3b91ce94-53c9-4e2e-976c-7b955099b888 · outbound

This paper cites World Model for Robot Learning: A Comprehensive Survey.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models World Model for Robot Learning: A Comprehensive Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.819932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.819932Z digest=sha256:cb6448a54378356b9c885f946bd8905dba9ec8bba9c1c8b3c35c74b9962a1382

Observation dcd1afdb-0f74-4bc3-9709-a984eff04d8b · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GAIA-1: A Generative World Model for Autonomous Driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.824533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.824533Z digest=sha256:04b9a9bef00d3c30133be290b64c90057ad1d14b3182675eb0a85f1fc4fa1972

Observation 5f92cb20-abac-4f4a-8b4b-aac060782734 · outbound

This paper cites PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.829126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.829126Z digest=sha256:7539a04fb11e55efc3a1ac379db4192a7cd6ac3ae62854fcd2cb3ec14f5c12c3

Observation 1fc97768-26d9-43a6-be44-ef0a5a4f0c8e · outbound

This paper cites Robots pre-train robots: Manipulation-centric robotic representation from large-scale robot datasets.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Robots pre-train robots: Manipulation-centric robotic representation from large-scale robot datasets

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:41:44.680657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.833447Z digest=sha256:d8135700ff095c7a23183bd8618705acd80520ff19a6ca6f29c09ba5e7cba9d7

Observation 7ea2f6d8-a26b-4621-8ca7-0e96105cb8e0 · outbound

This paper cites Contrastive Representation Regularization for Vision-Language-Action Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Contrastive Representation Regularization for Vision-Language-Action Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.838078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.838078Z digest=sha256:a306a2f14bb287277c34c7c3ebeb796fe32d1563a00823ae9f51aa00a9f5b087

Observation af3f1281-6c88-4867-b91b-fddb33a74a6e · outbound

This paper cites Predictive but Not Plannable: RC-aux for Latent World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Predictive but Not Plannable: RC-aux for Latent World Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.842384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.842384Z digest=sha256:f2a3db3a1112c321891cc1a8535d7169ef4df44a568e10bff9bfe35aeffaa288

Observation 2bb77745-38f2-4b85-b1d7-189a61b52d8a · outbound

This paper cites Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.846637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.846637Z digest=sha256:c68c503d709f3e98c6b4ed63059197b0d3cd1ff80846f76180d00547a15b67bf

Observation 51beead7-cf0f-4c5a-bc12-9285f117fa83 · outbound

This paper cites LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.851098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.851098Z digest=sha256:7f1ace5f40ea91142791b347d17b1069b3cf92dcd84fdb8daf030f0e9896cd8c

Observation aedfe381-ac0e-4ae4-9857-f3162503142e · outbound

This paper cites LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.855678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.855678Z digest=sha256:a770303a3683b7d629a48fa3332dfee95e6c9bef09301bd98341194a9c71d85f

Observation 643e5636-e55b-4e10-bad6-760151b930b4 · outbound

This paper cites V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.859391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.859391Z digest=sha256:8b405750d8fcc85f313a31e114586012ce40f210dd55bde985de637568c8c11b

Observation 67a071b4-b42c-4f8a-adcc-b42a81e085d4 · outbound

This paper cites Latent Geometry Beyond Search: Amortizing Planning in World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Latent Geometry Beyond Search: Amortizing Planning in World Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:41:44.410479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.863650Z digest=sha256:64000501ef6738324656d435d109f6953050d405c65d57b6fbf76db6f3d3ede6

Observation 10945e6d-0db1-408e-b76b-d182625c2ba1 · outbound

This paper cites LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.867336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.867336Z digest=sha256:c424775f21032951cbe36fb061d646bd2d4b184bbb168f42a7c5fae9a3b57dfc

Observation ec375ccf-128a-4bf5-a12b-5cc8b3c77bcf · outbound

This paper cites Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.871583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.871583Z digest=sha256:a466b62cc3d7080a5c35fd541ae900d0c7972e070215791aafb7638b80770fd8

Observation 2efa1da1-fa8e-4df0-a553-911bd1a2e13f · outbound

This paper cites OGBench: Benchmarking offline goal-conditioned RL.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models OGBench: Benchmarking offline goal-conditioned RL

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:41:44.668012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.876052Z digest=sha256:85b9d30302409f91499b083e8469daec304eb5150bbe80120a5998cd9d12fafb

Observation 8c808325-6709-4310-8b71-c8b8d37aec99 · outbound

This paper cites an unresolved cited work.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:41:44.655923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.880350Z digest=sha256:2814824d78f395efb29a0c27686728492dfe2aff071ebe7e9e3b4c5aa3f72d9b

Observation 5ab9f07f-3493-4110-82c5-0dec7157c14d · outbound

This paper cites GigaWorld-0: World models as data engine to empower embodied AI.arXiv preprint arXiv:2511.19861,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GigaWorld-0: World models as data engine to empower embodied AI.arXiv preprint arXiv:2511.19861,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.884472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.884472Z digest=sha256:b09314347a28ae70498032242017edd35e88b35d247aab5e234a1ade6aa400e6

Observation fa3a742b-7e23-4cc6-bc98-485f524eae8d · outbound

This paper cites GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:41:44.293651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.888041Z digest=sha256:693b79dda878a53285b85593176f9a41baa3d40db64110e7dca51dfd3f342cd0

Observation 100d3085-3228-4b68-934e-1b92b13bcb18 · outbound

This paper cites Beyond language modeling: An exploration of multimodal pretraining.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Beyond language modeling: An exploration of multimodal pretraining

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.891873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.891873Z digest=sha256:c4e7dec89546b5e584a14195ce02ffe27ad0c1456639318e3518313aea85be89

Observation b2b4d677-4f69-4894-89b2-530598783472 · outbound

This paper cites Open-world hand-object interaction video generation based on structure and contact-aware representation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Open-world hand-object interaction video generation based on structure and contact-aware representation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.895990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.895990Z digest=sha256:dbc1cff512a7d8f19ceca40cb30bc0a41cd520b6cd1dfcdd257627b79fb05239

Observation fd1f48ef-37a7-4918-ab92-6647f9a4f302 · outbound

This paper cites UniDriveDreamer: A single-stage multimodal world model for autonomous driving.arXiv preprint arXiv:2602.02002,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models UniDriveDreamer: A single-stage multimodal world model for autonomous driving.arXiv preprint arXiv:2602.02002,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.899958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.899958Z digest=sha256:f023eec754186f078893282944750f00ce5218787e87c49431be5c0b56f84426

Observation a6211a8a-2fbd-4454-aef9-41f3e0c4ef1e · outbound

This paper cites FlowVLA: Visual chain of thought-based motion reasoning for vision-language-action models.arXiv preprint arXiv:2508.18269,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models FlowVLA: Visual chain of thought-based motion reasoning for vision-language-action models.arXiv preprint arXiv:2508.18269,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.904159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.904159Z digest=sha256:2e22c019f378926d77c739650cdffc7c4377f32bb6d297c7e7d3465e0ea4edb2

Observation 4de75839-f51f-4d07-8fc8-2261a689b7ea · outbound

This paper cites DualCoT-VLA: Visual-linguistic chain of thought via parallel reasoning for vision-language-action models.arXiv preprint arXiv:2603.22280,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models DualCoT-VLA: Visual-linguistic chain of thought via parallel reasoning for vision-language-action models.arXiv preprint arXiv:2603.22280,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.908418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.908418Z digest=sha256:e1959ac273dc4c5e56482bad9eb0df9af7b2602ec1e25bf5e848f57a044425d7

Observation 45e2f42a-1cc4-4f11-9eae-04f23fe607f5 · outbound

This paper cites DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.912335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.912335Z digest=sha256:accf007ab170001c55950114d445476e5a67ea21b84f547e8795262e69152c2a

Observation b57a9edb-0419-45a3-ae8c-8411cb48279c · outbound

This paper cites Mastering Diverse Domains through World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Mastering Diverse Domains through World Models

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.815090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.815090Z digest=sha256:cf97e7eae96de39d0aa84d66c2e7c40aa2015e78ab3f3aecf32e1c7d9ee58be0

Observation cec43895-c873-4d31-8f07-e2c3850ce7af · outbound

This paper cites V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.793178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.793178Z digest=sha256:ee7bda334e3361b99539d4d5da8d23f272a06f294f0cbd7072d413d9a881fa93

Observation bfdf302f-39df-4e2f-a3ec-b7a8a32bc3f2 · outbound

This paper cites Why ai systems don’t learn and what to do about it: Lessons on autonomous learning from cognitive science.arXiv preprint arXiv:2603.15381,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Why ai systems don’t learn and what to do about it: Lessons on autonomous learning from cognitive science.arXiv preprint arXiv:2603.15381,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.797534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.797534Z digest=sha256:50bd296f87026c5b9fda7470c3ddddef6a112e0146feaef7d5738e6aeb2ee3da

Observation faa62ea0-a425-480d-a657-9e928ba511f5 · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.801307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.801307Z digest=sha256:b670b3666b9a5d47b4e1f0bbba15a0bd2ff45943bda923b344e39b53fbae179d

Pith citing papers

No inbound Pith citation observations are available.