Pith. sign in

Paper Citation Record · LEDGER

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model

As of 20 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2605.31234.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.31234 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T22:04:09.296855Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T01:34:31.334140Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact25
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 35ba7f4b-76ac-43b1-8817-eabe773a0b30 · outbound

This paper cites Latent Action Pretraining from Videos.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Latent Action Pretraining from Videos

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.862338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:ace3f69b77d19ef264a247b600942c1f5d44765634ddb5881673c00b43bccc5a

Observation 32aa5940-6bd1-4895-baeb-e3d3a16511a7 · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.880387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:c0ea206db9e9dc0e64e6a8732e7167d08091d1159445ff24c4d60c292fc6e17b

Observation 7798bbd5-02ca-4d48-a471-b698704aeccd · outbound

This paper cites J., and Lee, Y.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model J., and Lee, Y

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.857254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:fa91cdee9bc4a8af5420d9ce5fb0bae7352e0184adfdb670592983fb2adc9532

Observation 43bdcb58-c2db-4c92-8681-e5411df3c993 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:a79999a02468ad3d18d1aa14c9f5efe6dc025f52a7f8c13e56bea6a0dfcee43b

Observation 064d0ce0-74bc-4ff7-8b8e-abfaca8ea110 · outbound

This paper cites Kareer, D.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Kareer, D

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:eb47bdf1c7123074ad2543860e191a10bf1d3087bd266d2bfd3b42e935255480

Observation fd2b0dbb-915d-462e-a9a2-8558caa50aa8 · outbound

This paper cites arXiv preprint arXiv:2509.22199 (2025).

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model arXiv preprint arXiv:2509.22199 (2025)

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.860090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:d269c73610854829ceead8e99215219b80530b46956ef17d77ddbe8b8ab0594f

Observation 6a093d09-7fb5-4210-a04e-d5fdfc8e3699 · outbound

This paper cites Dexumi: Using human hand as the universal manipulation in- terface for dexterous manipulation.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Dexumi: Using human hand as the universal manipulation in- terface for dexterous manipulation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.864905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:9598a5326eadab17673558548866d82c084149f2694fff6564d1f1b9dd862f4e

Observation 42fc5a09-bc12-4092-b6b8-cc41e2272440 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:456a8f776f323cac1a9bfd7448afd51e246d6a5ddd7a85958ce0d1534c17b24d

Observation d194bc2e-f932-440a-a3fe-6bbec9073162 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.864485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:30f3fb65d6cff64afd87a6afbd4bfe044db81b38bd37f95f05af50057368630c

Observation a03984bb-8b93-466d-b361-620c6c29afe8 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model RT-1: Robotics Transformer for Real-World Control at Scale

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.876832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:e70d639fd0144845665d01805383fd407aabb278d47518128f87b3b7864d2298

Observation e01e983c-4254-4ae8-a537-c8858f3b617c · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.861914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:69a820565541d004acdd28d4c4ee75f52a335c35427bd8fa76b2507c3de0b8e1

Observation 1c0df870-7193-4d33-bc0d-30c4d0db2665 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:5833b05facb22b6e50fa7e3453d5d5fabec533cfb66561eb97c77194f411383f

Observation 93e23b68-3d6b-4038-a72b-1225fb19e32c · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Octo: An Open-Source Generalist Robot Policy

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.859283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:8236ac30a1bce3b27325b6dbe8c76fc199aaa58312348e7c196537b4f374f752

Observation dc4e691f-b7f0-493f-a413-67f31826560a · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.878008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:001a80c4bc8d508c18e46c7d5a32c2c33e9e6a6d28bbe3e2b5d2a26df4c451d4

Observation 0ea0dfb9-827f-4003-ba3f-9db05ebc6cad · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.869549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:c94ebd92e954568c4fbdb98714e0534d3d003c6fd1d547731b095aa553277011

Observation 5879ffed-fee5-4786-8183-47c259e5ac3f · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.872735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:aaab9472ed4a659c1c30ee3d6020805548473ca3664909e8327797e4e9b793bc

Observation 70717396-c379-4ad4-be3c-847929f6a063 · outbound

This paper cites RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.828518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:3ab51b65aa3321bcc7b84b147dc4bc27c7d9bce37c426ea00ae8bbe3a12ac4e6

Observation 36f83a56-0072-4c87-b74b-9317167cb482 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:b64450af280828e179db5feee57cfe62b29e1ecf0901408af5620815a8debf30

Observation 66491bf7-482b-402a-8085-bb8907fa9795 · outbound

This paper cites Xiong, Q.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Xiong, Q

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:670aa7c9abdac2bdc04e5d389c3f7adee37c0940ecad17adc3fae6b82331efbe

Observation ad161016-7bb7-4fe6-ad2b-bb91d39a2f9b · outbound

This paper cites Flow as the Cross-Domain Manipulation Interface.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Flow as the Cross-Domain Manipulation Interface

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.870093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:4c48ad60573ad22e79e05ff6597ce6a72c67f49759107638052d95392a8c91cc

Observation 053c5aec-d184-4b12-89d2-c6a7fffd8c09 · outbound

This paper cites Dream2flow: Bridging video generation and open-world manipulation with 3d object flow.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Dream2flow: Bridging video generation and open-world manipulation with 3d object flow

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.879323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:1ea41a109828048b25b040bc6509664f73223f070edc2b88a17a3ba33e5e5e2d

Observation a5149324-bddb-46b0-8e8f-b2ba96f0aca5 · outbound

This paper cites RoVi-Aug: Robot and Viewpoint Augmentation for Cross-Embodiment Robot Learning.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model RoVi-Aug: Robot and Viewpoint Augmentation for Cross-Embodiment Robot Learning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.875390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:10254e045023f943d9169a8bf0381bc0cd6418695091bdc01b9db4606c586354

Observation 0c48d54d-9fdb-4815-a340-8eb285135d4a · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:dbda4306a330e886a6b5583065f5d0c9df26d911a8ca18490a7e44cbde1b45ed

Observation 6f7be1f5-8fa4-4301-92af-9add02daba39 · outbound

This paper cites OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.853926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:7d683fc589d53333b58b7efe50157f66c022ed7cbdf9407974c426af4fb8765e

Observation d98d2698-de49-45ff-85cd-5f29bd1ba444 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:8984e8c235648bb667345610f22d672e46ee6d078948c5e551db191936e778bc

Observation dc173fc5-0de5-45e0-b9d2-7a24665d9ae0 · outbound

This paper cites RH20T: A Comprehensive Robotic Dataset for Learning Diverse Skills in One-Shot.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model RH20T: A Comprehensive Robotic Dataset for Learning Diverse Skills in One-Shot

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.866879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:42d82ed508102ebfe4f73dbc61a3f47adb240b9bddde4b21b453a6ffe32a989a

Observation 0528053c-90e5-4e9a-a3be-2f2609f06edc · outbound

This paper cites Human2robot: Learning robot actions from paired human-robot videos.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Human2robot: Learning robot actions from paired human-robot videos

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.840796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:31b22b06aa5b953d342f52780c2de5a55cc4693295f5d74f2a2a130b3cc655a5

Observation 3d1d54bc-cab0-4e2a-8457-312c3ba68754 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:812901fc465f8361f19815091fa8cb59c92a5eed73dc1a333ff91c392a58ac95

Observation bcd5b816-09a3-4b20-9c41-c3958f8bc118 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:58a48615e5ad83512174a889a105c1dc924df4105e28c1c1d2fda53895596eb8

Observation f5e5bddf-bb26-453c-aaaf-b6070511d466 · outbound

This paper cites James, Z.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model James, Z

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:e4c636f132ae0b83635dd0729eb13a19a3828f7de5231d5de968dac720046002

Observation b9d80d99-ee8b-42f0-8c7a-ff17d73a253a · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.874408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:17b79d269584dbb16ecf5e136db672c4ecf896c273b292537644e7aade54e361

Observation 2fdecf46-b87f-4a73-8443-9f2a24b1aaad · outbound

This paper cites Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.871892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:8f7b903e3410a439bbd661f4a1b3e0947d231a6922535f822379667b1b438a1c

Observation cd3b6cb3-e617-48f6-9a23-d167f338a5e0 · outbound

This paper cites Qwen3 Technical Report.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Qwen3 Technical Report

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.828901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:b9f85983e6eed25f70ac2b8aa83729862acc1f4ff0184d5e7322251f8b0ec527

Observation 3d7ee676-4f8c-4a4e-9f65-577cb22c9551 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:075a5377a6320d0f89023a0f29d15e72128a45da2b3ae7147baa193e1127c07c

Observation 13b58a9c-ff64-4a76-9aa0-11893ee77d72 · outbound

This paper cites Doersch, Y.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Doersch, Y

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:02f30556dc4c013b0ea4496bec546e2f04f2f459657bc55d78bd12416b896b7d

Observation 9c1cc4fc-6f01-4c0e-9c69-abcbfed54641 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:fec95b0a0222029664e84397a22d6265b526a786fd1fba765b7075d57a516a7c

Observation e63e87b2-d476-41d7-898f-25fd56ec8bec · outbound

This paper cites Embodied Hands: Modeling and Capturing Hands and Bodies Together.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Embodied Hands: Modeling and Capturing Hands and Bodies Together

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:10.842693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:bf8700cccc546fb514f4f8af228379b6a13dde5bfce4f2ccfa04b364299ad84e

Observation e9d5f882-8369-41c2-812e-47a56fd1ab28 · outbound

This paper cites Yadav and M.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Yadav and M

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:eedddde87a14b2fc9c50dda524d10beae57de9dfe53170e1758feb46046ab3fe

Observation aa901d56-61d9-4944-afe0-b4583f30dfaa · outbound

This paper cites Karamcheti, S.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Karamcheti, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:498e7b8e193c4ea261840342277e4f7e562bafd67763ff8bfdf4d867190d6560

Observation 47cfa813-6f33-4c28-a66f-e58eb02e9f15 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model DINOv2: Learning Robust Visual Features without Supervision

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.850895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:4230ea53c2859667c3b381d86f38a37d5aef535071084b180fa7848bb9636dda

Observation 2fa1af2e-1e3d-43e8-83c9-0860a7a943a2 · outbound

This paper cites an unresolved cited work.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T22:04:09.296855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:386f8227f85ed3349f2bb6f2485411ed975b50d3d34ddc30e02866273b24e409

Observation 4a3a4bea-dcba-45ba-a093-82660c00d795 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:46:10.834458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:04:09.296855Z digest=sha256:eb56e6f6d2eaa18984cd1d26378c2fd6ebd4b1c12ee526577c3eb92610ab6e60

Pith citing papers

Observation 074e1e22-2d11-4cd4-b40b-49822b0b9199 · inbound

Action QFormer: Structured Representation Shaping under Action Supervision in Vision-Language-Action Models cites this paper.

Action QFormer: Structured Representation Shaping under Action Supervision in Vision-Language-Action Models HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T01:34:31.334140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:34:31.334140Z digest=sha256:b42e63546a9438458a64357497f77766c9bf95843d2118c996ac331115f2694d