Pith. sign in

Paper Citation Record · LEDGER

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models

As of 12 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2608.06799.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06799 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:41:43.912335Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact2
  • verified fuzzy2
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c0ab036a-f408-45a3-84df-45276b015293 · outbound

This paper cites Ctrl-World: A Controllable Generative World Model for Robot Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Ctrl-World: A Controllable Generative World Model for Robot Manipulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.805937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.805937Z digest=sha256:a218b98e6239d789c568b9c7e6f391a88428401c91511832fffc3c4a0c9e999c

Observation ff021cfc-2268-44d6-ab40-4170ef71fb38 · outbound

This paper cites World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models World Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.810581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.810581Z digest=sha256:fcda3d3fc5dbd452b04b9e963c665861c3c40d047e8f6b6cc99509d541359b36

Observation 3b91ce94-53c9-4e2e-976c-7b955099b888 · outbound

This paper cites World Model for Robot Learning: A Comprehensive Survey.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models World Model for Robot Learning: A Comprehensive Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.819932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.819932Z digest=sha256:4b5f0c0e150a45db4b359ddd7d391376398cb7d066e16e07e040626d905c549e

Observation dcd1afdb-0f74-4bc3-9709-a984eff04d8b · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GAIA-1: A Generative World Model for Autonomous Driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.824533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.824533Z digest=sha256:742093327f0b4a6e5d96d2db39861fd490be76b469cb5086a7bde7b2812e3de5

Observation 5f92cb20-abac-4f4a-8b4b-aac060782734 · outbound

This paper cites PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.829126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.829126Z digest=sha256:3f944720d58a071cdc0eea48c95c52797dce4210437a62cade1936f3405a4747

Observation 1fc97768-26d9-43a6-be44-ef0a5a4f0c8e · outbound

This paper cites Robots pre-train robots: Manipulation-centric robotic representation from large-scale robot datasets.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Robots pre-train robots: Manipulation-centric robotic representation from large-scale robot datasets

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:41:44.680657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:41:43.833447Z digest=sha256:d4ccda6077c6841919266a8eef04c64ec40f8559e71ccd1acdd05cf5b39aa1d7

Observation 7ea2f6d8-a26b-4621-8ca7-0e96105cb8e0 · outbound

This paper cites Contrastive Representation Regularization for Vision-Language-Action Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Contrastive Representation Regularization for Vision-Language-Action Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.838078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.838078Z digest=sha256:4d756f2702e274285c6db5f387acdcc6a6579230f4448215a932f30648b52070

Observation af3f1281-6c88-4867-b91b-fddb33a74a6e · outbound

This paper cites Predictive but Not Plannable: RC-aux for Latent World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Predictive but Not Plannable: RC-aux for Latent World Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.842384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.842384Z digest=sha256:5df1509d6fcd29f302b5a93a2cc2eff213e1e48faf43fcf34e0b95a2488f12a7

Observation 2bb77745-38f2-4b85-b1d7-189a61b52d8a · outbound

This paper cites Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.846637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.846637Z digest=sha256:9fdf3ed2f52bee8ea1b5502c718037cd7f3c203b361ee4c2f9449adee77fecc1

Observation 51beead7-cf0f-4c5a-bc12-9285f117fa83 · outbound

This paper cites LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.851098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.851098Z digest=sha256:1a9ebbaa27b0ba04648771341fc3a7b40104d30e14ebc8a9416f4c90df77f05e

Observation aedfe381-ac0e-4ae4-9857-f3162503142e · outbound

This paper cites LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.855678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.855678Z digest=sha256:8917125f4fbafdedea15c1c49a739b084232fe44d74fe9c5676556d5687e6726

Observation 643e5636-e55b-4e10-bad6-760151b930b4 · outbound

This paper cites V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.859391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.859391Z digest=sha256:5e939072e4d2fc7121c235c2d747a2cffe2f972d046e14985f7295688060cf95

Observation 67a071b4-b42c-4f8a-adcc-b42a81e085d4 · outbound

This paper cites Latent Geometry Beyond Search: Amortizing Planning in World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Latent Geometry Beyond Search: Amortizing Planning in World Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:41:44.410479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:41:43.863650Z digest=sha256:130aec0ac9ee8a4269049e0de85dd98b34412d7cb9749a7f0e971759d3c6e742

Observation 10945e6d-0db1-408e-b76b-d182625c2ba1 · outbound

This paper cites LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.867336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.867336Z digest=sha256:07588c3ea75f727c6189732f08bb60bc325907812c024fc9e4331ac8f9d263ee

Observation ec375ccf-128a-4bf5-a12b-5cc8b3c77bcf · outbound

This paper cites Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.871583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.871583Z digest=sha256:e52ee5c3d7ec18da3d32baa9f48e3b5a5b5698ab6f2f17aa52e741b02c49b371

Observation 2efa1da1-fa8e-4df0-a553-911bd1a2e13f · outbound

This paper cites OGBench: Benchmarking offline goal-conditioned RL.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models OGBench: Benchmarking offline goal-conditioned RL

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:41:44.668012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:41:43.876052Z digest=sha256:639ea47cad8fc05eac7324f6d904a2dd8aa9854984df8c8acb1318de7a2511db

Observation 8c808325-6709-4310-8b71-c8b8d37aec99 · outbound

This paper cites an unresolved cited work.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:41:44.655923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:41:43.880350Z digest=sha256:0fb75ef164f16d71f7b778c1d9ee92d6e1365bbc1fcd5f81e6223f3dd003ef46

Observation 5ab9f07f-3493-4110-82c5-0dec7157c14d · outbound

This paper cites GigaWorld-0: World models as data engine to empower embodied AI.arXiv preprint arXiv:2511.19861,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GigaWorld-0: World models as data engine to empower embodied AI.arXiv preprint arXiv:2511.19861,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.884472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.884472Z digest=sha256:5dcd8f19ec092e39a50bb5b32fcb8d8dcf62bd01f71a7497b554a8da3ebe58c9

Observation fa3a742b-7e23-4cc6-bc98-485f524eae8d · outbound

This paper cites GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:41:44.293651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:41:43.888041Z digest=sha256:ad6eec285a41666edffbbed4cd5f445972b7d3e04fd2da5eb7eba859586ab3e7

Observation 100d3085-3228-4b68-934e-1b92b13bcb18 · outbound

This paper cites Beyond language modeling: An exploration of multimodal pretraining.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Beyond language modeling: An exploration of multimodal pretraining

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.891873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.891873Z digest=sha256:de63b98f62656a1848ae96471d6c7751b92412187b73d258943ad680fdb7943b

Observation b2b4d677-4f69-4894-89b2-530598783472 · outbound

This paper cites Open-world hand-object interaction video generation based on structure and contact-aware representation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Open-world hand-object interaction video generation based on structure and contact-aware representation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.895990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.895990Z digest=sha256:36a73a429bd08d72d37879aea9592e8ba20b3a86644c17c0c41bc87c4811ed24

Observation fd1f48ef-37a7-4918-ab92-6647f9a4f302 · outbound

This paper cites UniDriveDreamer: A single-stage multimodal world model for autonomous driving.arXiv preprint arXiv:2602.02002,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models UniDriveDreamer: A single-stage multimodal world model for autonomous driving.arXiv preprint arXiv:2602.02002,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.899958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.899958Z digest=sha256:9af7db83897efd2bcd1a243e3a87ce93fa3c91595618ec2d6334f5c300a30f1a

Observation a6211a8a-2fbd-4454-aef9-41f3e0c4ef1e · outbound

This paper cites FlowVLA: Visual chain of thought-based motion reasoning for vision-language-action models.arXiv preprint arXiv:2508.18269,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models FlowVLA: Visual chain of thought-based motion reasoning for vision-language-action models.arXiv preprint arXiv:2508.18269,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.904159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.904159Z digest=sha256:538a8578e0c6d0fc90294d41de6d2e98f2d2ea4e4491da90c32fbda84f9f93a5

Observation 4de75839-f51f-4d07-8fc8-2261a689b7ea · outbound

This paper cites DualCoT-VLA: Visual-linguistic chain of thought via parallel reasoning for vision-language-action models.arXiv preprint arXiv:2603.22280,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models DualCoT-VLA: Visual-linguistic chain of thought via parallel reasoning for vision-language-action models.arXiv preprint arXiv:2603.22280,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.908418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.908418Z digest=sha256:b88e2fbdf71a1cc6a0b816cf9c286492a796bb8916ce5aa4fcd5d5f48d33a385

Observation 45e2f42a-1cc4-4f11-9eae-04f23fe607f5 · outbound

This paper cites DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.912335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.912335Z digest=sha256:bfe249a7ffd4cfa91c3c049ee50d34535ac7a3614729dfc6d04b385e4ffdd9d6

Observation b57a9edb-0419-45a3-ae8c-8411cb48279c · outbound

This paper cites Mastering Diverse Domains through World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Mastering Diverse Domains through World Models

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.815090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.815090Z digest=sha256:5d18d348b368d616d3e6342450e2263abb5cb525b538598ffc801356d29819eb

Observation cec43895-c873-4d31-8f07-e2c3850ce7af · outbound

This paper cites V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.793178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.793178Z digest=sha256:894212e667cfc9b60f8c6acd137bfe4c6c4b6b58ddd436beef7773c51bf5fa12

Observation bfdf302f-39df-4e2f-a3ec-b7a8a32bc3f2 · outbound

This paper cites Why ai systems don’t learn and what to do about it: Lessons on autonomous learning from cognitive science.arXiv preprint arXiv:2603.15381,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Why ai systems don’t learn and what to do about it: Lessons on autonomous learning from cognitive science.arXiv preprint arXiv:2603.15381,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.797534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.797534Z digest=sha256:7dc93e75da459f54e454e1e25b9057dd6c996d72f85be9f073f743f3e595af9e

Observation faa62ea0-a425-480d-a657-9e928ba511f5 · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.801307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.801307Z digest=sha256:4506d65d639b7292efe3e9ba787d948c2e9a5bbea3209636e9642b7559ae17aa

Pith citing papers

No inbound Pith citation observations are available.