Pith. sign in

Paper Citation Record · LEDGER

IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2411.00785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.00785 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:33:12.577205Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T21:00:09.783700Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 547818d4-8ee6-4805-ad5e-831fc3e2775a · inbound

VideoDPO: Omni-Preference Alignment for Video Diffusion Generation cites this paper.

VideoDPO: Omni-Preference Alignment for Video Diffusion Generation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T12:29:09.438211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:29:09.438211Z digest=sha256:c6d4fd8b94e3c078c3a5cabddbab47b18924c7a3108033b17e087449e5ac275a

Observation 9b581b68-ccdd-45de-b531-0391d1bd19c7 · inbound

Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations cites this paper.

Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 93

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T18:38:11.313568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-12T18:38:11.110166Z digest=sha256:f19e50be828892d497ab29ed6340e995d22d0ae910cbb3a2f731db38e6a620cf

Observation f78bd75c-6e8c-4692-8777-791cbb8c8c34 · inbound

UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent cites this paper.

UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T22:13:14.887809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:13:14.887809Z digest=sha256:957db3725c92ef74c15c55d8887a2fd9f3c56056e60ccf794e309abe3ce42175

Observation 767b7ea1-b712-411a-9dc9-a8626f3ec87f · inbound

Latent Action Learning Requires Supervision in the Presence of Distractors cites this paper.

Latent Action Learning Requires Supervision in the Presence of Distractors IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T19:18:44.893686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:18:44.893686Z digest=sha256:bba08a2fab197c4bd6cc27c658c1627d654211dce5f0fb0495123a6ae65e767d

Observation 91ebf4bc-00de-4437-80f8-de0c067b9d74 · inbound

UniVLA: Learning to Act Anywhere with Task-centric Latent Actions cites this paper.

UniVLA: Learning to Act Anywhere with Task-centric Latent Actions IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T15:28:06.967896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T15:28:06.883492Z digest=sha256:64d9f805a53e27b15193ea28743fe4b9edfb8d752b6501b9c0ec83b0d374557d

Observation 606d6697-a9b2-4e10-8309-64266fbc35a7 · inbound

Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation cites this paper.

Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:33:12.577205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:33:12.577205Z digest=sha256:0a8d870532df0eb941cdf25a5cf0f68de9f28eaf1e779c2c70e7267bc273b76a

Observation 87d28d03-a0ad-4bca-a374-5a7b333b7764 · inbound

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning cites this paper.

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:33.139916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:56:33.139916Z digest=sha256:3dc2f3374d93ebc5e78ab4b7eaf491016c6dab32b7c1ec278e5ed58261a41384

Observation 6d8aa1ed-3ea4-4adc-9999-2e46881603c9 · inbound

WorldEval: World Model as Real-World Robot Policies Evaluator cites this paper.

WorldEval: World Model as Real-World Robot Policies Evaluator IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:41.835024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:41.835024Z digest=sha256:d6ab064c394fa7a1dec174701b52dd73d96dd1472b9571802fd1f0faa4880f10

Observation 794491b1-9a61-44bd-8399-e6d5650962a6 · inbound

Playing with Transformer at 30+ FPS via Next-Frame Diffusion cites this paper.

Playing with Transformer at 30+ FPS via Next-Frame Diffusion IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:49:42.207829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:49:42.207829Z digest=sha256:2f00af38872f5c6f002acac54e387439958177e16eaeee37b92ae3f3f61f1c8f

Observation 7cbd0345-3ac5-4cbe-85e4-6bda7755a4c8 · inbound

GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation cites this paper.

GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:50:29.231047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-25T07:48:33.832017Z digest=sha256:e46517f88898d2d4179956cd97de43719395c9e3ab105f314e681d74ccf25710

Observation 4ddca0e3-d259-40f1-a1f5-800e9a5cb0ce · inbound

VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models cites this paper.

VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T19:10:27.083480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:10:27.083480Z digest=sha256:aa22387c721e09fafd52be626dbae9071362176430c4e031baa642de3622c705

Observation 3827fc8f-d211-4da5-b32d-9e48b3d2b97c · inbound

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers cites this paper.

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:07:22.759775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:07:22.759775Z digest=sha256:4e318d1f771a7b5042ee1a82bf504df3fd00aeebc35aa95e1d2766862c955d06

Observation 368c1dee-66e7-4e08-a7c8-d6b3dfddfcbb · inbound

Is Diversity All You Need for Scalable Robotic Manipulation? cites this paper.

Is Diversity All You Need for Scalable Robotic Manipulation? IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T19:17:10.758757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:17:10.758757Z digest=sha256:7a7f9735341165ba17aee60ef2ce409323f0ef57ced98ed8c1d128ab5c7f9a56

Observation dd2ef776-c9c6-459e-ad1a-8a7c1a38296a · inbound

villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models cites this paper.

villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T21:52:02.980898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T21:52:02.893886Z digest=sha256:559e057e7e82aa97b9deb2d64037c98876f71c93dba6cadd8add402c40d70807

Observation 6efd5a22-aa1e-4ed5-92a7-f8b23806f8cd · inbound

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation cites this paper.

PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T13:24:44.036005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:24:44.036005Z digest=sha256:739e54f58935b0662bde6dd17aa90836c7ccc2ccdf9ea243bdd9dcd46c2ec12d

Observation 30b947d9-96c0-4a33-b2c0-49c890709df5 · inbound

DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos cites this paper.

DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:02:34.146487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T17:02:33.997887Z digest=sha256:c665bdc23cfdaee369380bb548d3692c69d8fb55c231eaa3ae2329f21428aedb

Observation ad1e5935-e2cc-422b-9efb-6a396a4b0f5c · inbound

UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models cites this paper.

UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:20:17.665714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T20:18:31.988002Z digest=sha256:91b9213828f23fd96e0f585b58bb3db4965f198b5ff8e0eac4c5aa6a4810038c

Observation 64c0227a-a4b0-48ff-b393-3a78a56503d8 · inbound

Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training cites this paper.

Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:36:07.668999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-09T21:26:26.540403Z digest=sha256:31ec59200685c44b4043d6b0b3dc574ae73c5e3cffb38685988721c5c1278739

Observation ba7bbf18-1bc3-4ac1-84f5-bfff8fdfe94b · inbound

GazeVLA: Learning Human Intention for Robotic Manipulation cites this paper.

GazeVLA: Learning Human Intention for Robotic Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:36:12.527612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T11:37:30.784513Z digest=sha256:8386c091ef6bd6773dca576bab1c3197fce835ba556b4cdad3dc61e09d13c1d0

Observation 46bce8bf-d9c5-4aaa-aebb-c581e90d3716 · inbound

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation cites this paper.

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:26:11.673687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T02:51:27.662262Z digest=sha256:274213606adea8cd03042b47148a4f8c69ff4c4fb2258e5bcc165b107ba2ace6

Observation 8de5b719-e340-4891-9f79-461ecb42440f · inbound

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation cites this paper.

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:21:29.238755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T11:16:58.104663Z digest=sha256:8d1fff83c645de30014c13b861e1fa83a59c7290e2bea96bd80bff8cfdd77a2a

Observation 965afae3-47c1-41c4-a357-a712ee39d9ae · inbound

From Pixels to Tokens: A Systematic Study of Latent Action Supervision for Vision-Language-Action Models cites this paper.

From Pixels to Tokens: A Systematic Study of Latent Action Supervision for Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:16:09.928505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-08T17:43:00.453164Z digest=sha256:fb6dc4f5cccb589a1426ec7a2ecac41079aba7029a81b32f8751253ce03c0c55

Observation 7331027b-87d9-40cd-9f0f-fd2ab52caec9 · inbound

VLA-GSE: Boosting Parameter-Efficient Fine-Tuning in VLA with Generalized and Specialized Experts cites this paper.

VLA-GSE: Boosting Parameter-Efficient Fine-Tuning in VLA with Generalized and Specialized Experts IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:26:10.252214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-08T09:11:21.715023Z digest=sha256:5f27c65968353e42ad2087d575a6e2247301dcd1dc25641e22feb44b8095dd21

Observation 5594f8e5-fee1-41e7-86b4-ce891a6aa4e8 · inbound

ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models cites this paper.

ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:26:27.062999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T04:14:54.885244Z digest=sha256:e3d3d6bab57677edcba1d47cf809124d0bb8a5d04d6971197551b35a9cf851ee

Observation bf2bea7b-7d9b-48df-aab2-66738c7dc589 · inbound

ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models cites this paper.

ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:17:59.505091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-14T21:14:56.501485Z digest=sha256:5e48d92a6f622111149ca65742a1ffeda485fa9e3d060edac4fd1f73abfd747e

Observation 3efa652d-3532-4adf-8404-f23d6b9d4d17 · inbound

RotVLA: Rotational Latent Action for Vision-Language-Action Model cites this paper.

RotVLA: Rotational Latent Action for Vision-Language-Action Model IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:49:23.120637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-14T17:48:06.734816Z digest=sha256:3021e57ba37b723b2e0a3f0130cc6577a4cced41d9833993428817580c77065a

Observation ec9a79a5-9e97-473b-bd2c-63a9685361e0 · inbound

DiLA: Disentangled Latent Action World Models cites this paper.

DiLA: Disentangled Latent Action World Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T19:38:56.388138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T19:35:37.527479Z digest=sha256:5ed587e7f903204757d8ce359fb1f02bcc08fac8c8d063c07164d9da326b9573

Observation 3163ce5d-1ad9-47d0-966b-5d5c7d73387d · inbound

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model cites this paper.

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 3

Resolution
malformed identifier
arxiv_id, observed 2026-05-19T19:37:43.756786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T19:36:17.173279Z digest=sha256:1a0f671ccc42880c36231dd1aa5e81ed4cb1c77a0a625092fa64594060070ce6

Observation e9adc3f5-8256-4c5d-b50d-54b231aa71d3 · inbound

UAM: A Dual-Stream Perspective on Forgetting in VLA Training cites this paper.

UAM: A Dual-Stream Perspective on Forgetting in VLA Training IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:28:55.099978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T19:24:57.339949Z digest=sha256:05d5ff05916d4d8f17179b12a99f1f2e9af941783a87609912d6ffdfc88fa00f

Observation 32c9b733-a34b-4e5d-bc26-1b09967633af · inbound

Why Latent Actions Fail, and How to Prevent It cites this paper.

Why Latent Actions Fail, and How to Prevent It IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:34:05.330376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T08:33:02.745886Z digest=sha256:103cbc3fb25e29f388d190289fcf3fa456ac57122e218ef5d786783c2e3752a7

Observation 172dd323-3d20-42be-9cab-eb70a0bdf684 · inbound

From Human Videos to Robot Manipulation: A Survey on Scalable Vision-Language-Action Learning with Human-Centric Data cites this paper.

From Human Videos to Robot Manipulation: A Survey on Scalable Vision-Language-Action Learning with Human-Centric Data IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:55:00.167016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T18:53:07.734871Z digest=sha256:7fd8019411196a29381dbeb0211b38bf4b9be0dfd40bd7c0e18637d9d1819c16

Observation 8344fbfe-375e-40f9-94ba-00c48daaa38b · inbound

CLAW: Learning Continuous Latent Action World Models via Adversarial Latent Regularization cites this paper.

CLAW: Learning Continuous Latent Action World Models via Adversarial Latent Regularization IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:36:29.583480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T09:55:00.402411Z digest=sha256:cee2fcf6202a133a9cffdf47c2da671d862fc061629bfbc18732f18a75d997cf

Observation 5c0ca728-7f52-46cb-9560-31f73155c281 · inbound

LARA: Latent Action Representation Alignment for Vision-Language-Action Models cites this paper.

LARA: Latent Action Representation Alignment for Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:47:09.499733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T22:25:17.522240Z digest=sha256:edee9ca41b2bcabc458c5fa6f9fa2829ac67b95d6e6ce33d838181998a07842c

Observation 77d76d54-83c7-4ceb-9f3a-5e1da62f7a68 · inbound

LARA: Latent Action Representation Alignment for Vision-Language-Action Models cites this paper.

LARA: Latent Action Representation Alignment for Vision-Language-Action Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:35:34.492973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-07-01T07:17:42.045939Z digest=sha256:d9742b434cd4a3bb33d44a7fd451f81c6c683eb05c4065d005390002471563ba

Observation 92fadd4d-7535-4bbc-a91d-d257a94dbd83 · inbound

Motion-Focused Latent Action Enables Cross-Embodiment VLA Training from Human EgoVideos cites this paper.

Motion-Focused Latent Action Enables Cross-Embodiment VLA Training from Human EgoVideos IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:39:04.235028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T21:53:27.323112Z digest=sha256:26b6cccfd033ece00f4c24fb5f3761636beb2ba6fe88bc579f96ffe0fa1c5707

Observation f2cded77-958a-4ab6-a065-0eeb889d2748 · inbound

Motion-Focused Latent Action Enables Cross-Embodiment VLA Training from Human EgoVideos cites this paper.

Motion-Focused Latent Action Enables Cross-Embodiment VLA Training from Human EgoVideos IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:49:01.900671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-03T23:44:35.238493Z digest=sha256:3b00a5d91cdaa96f5b83c4651bbf48d0ee5e6cca947e8c37bafe9de70f28bed1

Observation c276495b-7d4f-4f51-a882-7d22f3eeb120 · inbound

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models cites this paper.

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:49:38.568483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T14:07:31.139419Z digest=sha256:dee5893fcba95ca3ff2ab61d4e2ab3077f0372b7e575cf1fd8770fd9e508b1e4

Observation ef32b57f-519a-4253-b24c-a2b511622482 · inbound

Learning Action Priors for Cross-embodiment Robot Manipulation cites this paper.

Learning Action Priors for Cross-embodiment Robot Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.785386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:fd57aee97e13698548df8ae355e88f7a4a0cd2c03c47c3841099f309da8c8c7f

Observation db4fb4c8-8d36-4d37-864c-cb06c02743e2 · inbound

Causally Debiased Latent Action Model for Embodied Action Conditioned World Models cites this paper.

Causally Debiased Latent Action Model for Embodied Action Conditioned World Models IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-13T04:50:59.096166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T04:50:59.096166Z digest=sha256:e676d75919ad488ea571a79797fca3963259646e6335869f7481a8a0d81c1bb1

Observation 9d99e890-b6a8-4652-bca8-e6f8c2abf07d · inbound

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control cites this paper.

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T11:27:05.835624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:27:05.835624Z digest=sha256:cdd5e8dfc8d00a12011360d4cdb89759dfbe630afe9547a04fbaf29c19de04c3

Observation 1756cd0f-e81c-4df8-a802-d10edc9e8d97 · inbound

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control cites this paper.

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T03:23:48.964275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:23:48.964275Z digest=sha256:a4cbbc35b81a0c29868b073ceab8f56af4f077f07d0923bb9849109bbcb521a0

Observation 58573bcf-2373-411b-884b-74b3d872dc1c · inbound

ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow cites this paper.

ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T09:42:43.644737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T09:42:43.644737Z digest=sha256:b91385ce6a2a14b234c9877fd302899a2d043fe430ddcfa9ba6e8172ce924fa5

Observation 720ef27e-345d-4bce-b680-ae4155102304 · inbound

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer cites this paper.

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 216

Resolution
unresolved
no resolver link, observed 2026-07-31T08:51:28.134244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T08:51:28.134244Z digest=sha256:322950f73cb425aa8e634e1f16bb1ba384153608bb0ebf74e08276ea9eb3dec8

Observation dc6f7df7-6c3e-44a2-8987-523be07b7926 · inbound

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer cites this paper.

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 197

Resolution
unresolved
no resolver link, observed 2026-08-04T01:23:09.888871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:23:09.888871Z digest=sha256:8a6cf81724d918c7fb75bc0dc152914c1e2d9bd62c03318cd5077da973c51910