Pith. sign in

Paper Citation Record · LEDGER

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

As of 4 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2604.10677.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.10677 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:36:23.197843Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:54:20.104031Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact7
  • verified fuzzy43
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3941fef2-64f0-4bf0-998a-d21f116c088b · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Open x-embodiment: Robotic learning datasets and rt-x models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.480489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:f2c4380653d08f7c7b9f436da87b08f6b3acdb5aa7e746eec5ed5f72900bf880

Observation 792b1e76-3bab-4a37-bf0e-e6655d513d82 · outbound

This paper cites Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:03.890345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:19ecf6959ff153a28b82b599c8ded9e81e5f37bf58bf030692dbeaf5d5015a25

Observation f755dbfb-8825-42cf-b91a-7be92c29781b · outbound

This paper cites Airexo-2: Scaling up generalizable robotic imitation learning with low-cost exoskeletons.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Airexo-2: Scaling up generalizable robotic imitation learning with low-cost exoskeletons

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.512216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:86d8f3b4390f2391b0d757263639e9849d1a39eea2e216009e9f5f17f9bde39d

Observation 5fe3b5b5-122b-44a9-8128-6e253026f817 · outbound

This paper cites Data scaling laws in imitation learning for robotic manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Data scaling laws in imitation learning for robotic manipulation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.528984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:3cb85630581d4e708d29797670e845e3473a0136159b9974da5546de871b474c

Observation d9980239-9177-44d7-8fb3-20439cd7e729 · outbound

This paper cites π 0: A vision-language-action flow model for general robot control.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment π 0: A vision-language-action flow model for general robot control

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.518325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:dec2049fee2d7ef43b8e30c8afd88dc604cc0b5a61656b957d61e0ee775d7e92

Observation a7850250-86f2-4d2b-9f43-fd4148ace0c9 · outbound

This paper cites Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.546019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:4925c9fb9ce1a07944f2665b80d8d576ef845c088324b672721118910a8b8764

Observation cb228280-0514-4add-83f1-c7a4f20ea4af · outbound

This paper cites Robomind: Benchmark on multi-embodiment intelli- gence normative data for robot manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Robomind: Benchmark on multi-embodiment intelli- gence normative data for robot manipulation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.506554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:fa80831fe791c322210549b019546bdd99b43a3fe34165272d490a48c7d5b411

Observation 94be8688-51ee-4e70-b8b6-688e3593ae44 · outbound

This paper cites Droid: A large-scale in-the-wild robot manip- ulation dataset.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Droid: A large-scale in-the-wild robot manip- ulation dataset

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.485256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:bc3cbce39173f049a3be1835953398f214629432ce58b2fb2d3fd80846251458

Observation 2bf4a8b2-11fa-4748-8ee9-c7a1ab4de918 · outbound

This paper cites Taco: Benchmarking generalizable bimanual tool- action-object understanding.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Taco: Benchmarking generalizable bimanual tool- action-object understanding

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.497184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:1bf365dab632b3600b0e81a954ca6403332733ba4afe685a235332dfce2a1c4b

Observation 4304229c-8689-4f65-8258-96579398830d · outbound

This paper cites Oakink2: A dataset of bimanual hands-object manipulation in complex task completion.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Oakink2: A dataset of bimanual hands-object manipulation in complex task completion

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.475838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:573be66471c99c735e3216b01b4f5aeab7f47f3d5506b015b2cc2478ca3f7480

Observation 0171e4f3-3380-430e-b78a-ffa8205dccf9 · outbound

This paper cites OakInk: A large-scale knowledge repository for understanding hand-object interaction.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment OakInk: A large-scale knowledge repository for understanding hand-object interaction

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.522463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:0d222b7dae3f872dd4e1deb9bdb56d387caa1cded8956d391ec66262be71403e

Observation 11eeed79-4e1b-413c-8e14-6bc0794f5365 · outbound

This paper cites DexYCB: A benchmark for capturing hand grasping of objects.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment DexYCB: A benchmark for capturing hand grasping of objects

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.449811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:2d1d8475a10b6bb5c57a6fa6c3cb5e826bd6dd3e6721329ab8a9e50216465a92

Observation 8f8b08b9-855b-4629-8805-9cf77a4aa6d6 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Ego4d: Around the world in 3,000 hours of egocentric video

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.514877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:a10eb48a9efeeaf229baae7e719f66c5ed89c58eae31c0aff7e5396c7bedc520

Observation a9a8b939-9bf8-483a-b04d-bdb494dfd229 · outbound

This paper cites H2r: A human-to-robot data augmentation for robot pre- training from videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment H2r: A human-to-robot data augmentation for robot pre- training from videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:03.901752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8150ebf9e67805795157229c4aa2840043733e32af9ec0c7313409d6b4ace8b6

Observation 416a1d5b-d243-4075-ad91-716ec8917329 · outbound

This paper cites Masquerade: Learning from In-the-wild Human Videos using Data-Editing.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Masquerade: Learning from In-the-wild Human Videos using Data-Editing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-29T02:04:55.330788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:e0a82b1a59ba6344b559822feddbc6170b9cec166b7d5b7d47cb2d60dde9e5d3

Observation 512a1e3b-bfe6-46eb-81da-fa7bf6f290e4 · outbound

This paper cites Phantom: Training robots without robots using only human videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Phantom: Training robots without robots using only human videos

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.500107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:c81949473894af65ecd1d268350434140dcc4d9051a27ed3f174e7087960b81c

Observation a850c922-64b2-4363-9a20-9a86d2bb9819 · outbound

This paper cites AR2-D2: training a robot without a robot.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment AR2-D2: training a robot without a robot

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.542656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:3758e1bc3eacb7532244021a41ec07ea7a50baa729b2f3227c2d38c7e0b565a5

Observation 5d0db213-e176-4815-95a0-571f29016792 · outbound

This paper cites Emergence of human to robot transfer in vision-language-action models.arXiv preprint arXiv:2512.22414.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Emergence of human to robot transfer in vision-language-action models.arXiv preprint arXiv:2512.22414

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:03.910466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:cb674f94399faaf71375379d2f660639af5767eb202b9524ffeca0969ab878bf

Observation 2aee3763-1a2e-42f2-9c15-65bfa3df68b8 · outbound

This paper cites Egomimic: Scaling imitation learning via egocentric video.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Egomimic: Scaling imitation learning via egocentric video

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.478992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:7dc656d94728724d8f160dc5db04c5e0cc99e8d18f7d7f732cb1997bb6589a4d

Observation 256378bd-56c2-4e95-b54e-0812c74afaee · outbound

This paper cites Humanoid policy ˜ human policy.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Humanoid policy ˜ human policy

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.443304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:0fc5f93cdee35a0d0167f95b2f16972fdec3d53ed2f989c554be25cafd199164

Observation 2f849ef2-1c7a-4974-be45-c6cf61e388d2 · outbound

This paper cites EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:32:58.962571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:24772bab65aa21562b2937b2b678b0874fc8fd8327ea7d7d8ca5e68febd21768

Observation d29a1cae-e256-4c61-bddd-c1cc9eec44c9 · outbound

This paper cites Univla: Learning to act anywhere with task-centric latent actions.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Univla: Learning to act anywhere with task-centric latent actions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.456459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:c7273621ab750b59881e323acd1e788b6517a39aa2031a303bd59fdde1622819

Observation afa1b9c1-131f-4115-b3a1-a9e909790b3f · outbound

This paper cites Latent action pretraining from videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Latent action pretraining from videos

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.494175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:6fd85ca29c0457e6ecb931bf1b40ec6740d093f7346e3e373814f10dd78d880b

Observation 29ffa3af-ec0c-4ecb-82ea-0e6b8f377755 · outbound

This paper cites Moto: Latent motion token as the bridging language for learning robot manipulation from videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Moto: Latent motion token as the bridging language for learning robot manipulation from videos

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.531756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:d8d219e02dc9e12dd9ee832b1eb38bf3ffabc386e147e10d5449e96aab9b5b47

Observation 39e11551-b418-45cd-a60d-932d263b3fd9 · outbound

This paper cites Mimicplay: Long-horizon imitation learning by watching human play.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Mimicplay: Long-horizon imitation learning by watching human play

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.549326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8f9c58f11c60303ed849645c2718c69c6c4cab806e059a54f76b63658a68245e

Observation 4a09ee17-e49a-4193-95b5-a16e33e2b548 · outbound

This paper cites ViViDex: Learning vision-based dexterous manipulation from human videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment ViViDex: Learning vision-based dexterous manipulation from human videos

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.574822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:7bbcc579523afd85839aaa25a65cf94f7f0cfea122f6a6ea4d30ec30d86b304b

Observation 3b079b06-8263-496f-86de-7eaa3a251f79 · outbound

This paper cites Vidbot: Learning generalizable 3d actions from in-the-wild 2d human videos for zero-shot robotic manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Vidbot: Learning generalizable 3d actions from in-the-wild 2d human videos for zero-shot robotic manipulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.463957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:9afc43cff8436735d1feb914f2875cddfd0713af43dffa726393f955319c739f

Observation 568091b5-4891-4ec5-8f4b-3b8733d678f2 · outbound

This paper cites Zeromimic: Distilling robotic manipulation skills from web videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Zeromimic: Distilling robotic manipulation skills from web videos

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.459542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:554a1576a4b1103567d339f4125be8a76c096bb0bea11cd9065ab3ad2410c4c5

Observation 8f02b45c-42bb-496e-a423-efab5307458a · outbound

This paper cites Affordances from human videos as a versatile representation for robotics.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Affordances from human videos as a versatile representation for robotics

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.567498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:c0d6da2babed8040edaa79753821974405c46fca0cc24c2e7d4e2b2e44bb38b9

Observation 80f1abf7-ec1b-4e46-a41d-9742046cee8c · outbound

This paper cites DINOv3.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment DINOv3

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T10:11:03.937349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:9e5645733c68b9f4698418dff761c2e412cdb2dfd0265b6d29dd037f5761037e

Observation 56d46e99-0e1a-42bc-81f9-3f62040197bf · outbound

This paper cites X-diffusion: Training diffusion policies on cross- embodiment human demonstrations.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment X-diffusion: Training diffusion policies on cross- embodiment human demonstrations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.477914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:647f2eb9c475752ae3d1e2ebdf8bd164b3d385d944ba4c1f0d61800152973691

Observation 6669ea4f-a141-4ff5-8fcc-876a461a2e87 · outbound

This paper cites Deep imitation learning for complex manipulation tasks from virtual reality teleoperation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Deep imitation learning for complex manipulation tasks from virtual reality teleoperation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.439811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:af3afd72d705854fe338ac336e23ab9bd06040e566a8bf05c561e6119e876106

Observation 45ace35f-cb12-457e-840d-f2c60307588d · outbound

This paper cites The surprising effectiveness of representation learning for visual imitation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment The surprising effectiveness of representation learning for visual imitation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.453909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:7d01c6cc7c7b5f20999928fbe57018adeb23815df2b3c7d26d85ad370bbda216

Observation 7da78a45-f978-4693-aca9-0fd5eac66617 · outbound

This paper cites R3M: A universal visual representation for robot manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment R3M: A universal visual representation for robot manipulation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.552058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8214c78fbfe3fbf4bf67d4e45be3ca98ffcd2e23a66a98320d4ab5f32cbb9689

Observation 4472963e-0834-42ba-8379-8c938de6694e · outbound

This paper cites LIV: language-image representations and rewards for robotic control.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment LIV: language-image representations and rewards for robotic control

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.482309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:07772740a61f4043159c3dcb57eb12852d77a65dae827d8611067d00d79b02db

Observation a0538adc-f5b6-46dd-87c3-663021737fe6 · outbound

This paper cites VIP: towards universal visual reward and representation via value-implicit pre-training.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment VIP: towards universal visual reward and representation via value-implicit pre-training

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.557468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:4b66d5c60ee82840a63a2c14ae4c1b744c4897237c848a11f10f00b71479e851

Observation 1f353144-343e-4793-8e93-6a3ff7102fe8 · outbound

This paper cites Real-world robot learning with masked visual pre-training.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Real-world robot learning with masked visual pre-training

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.564248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:fcacfb9e6b7ba91c94d22d19c36220fb1022070ba3058e7985e9ee7616c463a2

Observation a9ed8a37-057d-4d90-9249-914e5b34661d · outbound

This paper cites Learning transferable visual models from natural language supervision.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Learning transferable visual models from natural language supervision

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.538428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8e662f97f785b9e0ea05ec4d6041ebeb1ad7f97e435471e493f6d833ac26d604

Observation 1c059970-cac1-4566-865a-7de8613e92e3 · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Dinov2: Learning robust visual features without supervision

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.560976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:ade3cd252cd6d700cad445283eb5d6e9310afced2fc0e6534408c764b9c710a0

Observation bf77b21c-ffbf-4a65-a6d0-e78aa8beff10 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment High-resolution image synthesis with latent diffusion models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.488396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:0ac6cdc404901dbb7e0fc9351b308bf3406ba5a0519734bc93caaec65caaf8d0

Observation 09a69b01-ceae-4a76-8a2a-7ea62db034bb · outbound

This paper cites Cage: Causal attention enables data-efficient generalizable robotic manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Cage: Causal attention enables data-efficient generalizable robotic manipulation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.472057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:af1ba45b23befa4de79f655165e2867db45b62884b5063436a421efb80e7f8a1

Observation b9c69c93-7e6b-4b96-8598-eb8a880bf777 · outbound

This paper cites Theia: Distilling diverse vision foundation models for robot learning.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Theia: Distilling diverse vision foundation models for robot learning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.491140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8bbf15ff376508f136a3014540f8affe833dad174f470071728a9303a3568848

Observation a5fcab8a-e620-4248-ab31-a0e9fc3df029 · outbound

This paper cites Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-09T03:07:11.470849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:ca72942f40343090122ebde4666a3a31215aea0b9aa0b528721fd06925b851f3

Observation d94e07cb-f2ed-4d53-a322-e97e882a07c1 · outbound

This paper cites Gnfactor: Multi-task real robot learning with general- izable neural feature fields.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Gnfactor: Multi-task real robot learning with general- izable neural feature fields

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.509537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:907ced6fa72fb270e9e25fac49d52a96167291c4bc3ae539e5ff1f6f132836a4

Observation 55c47b8f-f327-4d82-8dbd-4ff6f7d979c1 · outbound

This paper cites SAM-E: leveraging visual foundation model with sequence imitation for embodied manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment SAM-E: leveraging visual foundation model with sequence imitation for embodied manipulation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.453158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:49596faca461f7b1b354949822735f44b322860bf55c6b356161e0ce4de4ecd7

Observation 353280cd-a49c-4955-a82b-72ccf2ed27ba · outbound

This paper cites Spawnnet: Learning generalizable visuomotor skills from pre-trained network.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Spawnnet: Learning generalizable visuomotor skills from pre-trained network

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.571372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:eb4d0c30a39502f87a79be05b955fa7aa182ffe21fa2dfa197d643bb2f912bde

Observation 8ac2a8b8-7954-4970-899a-8655c09f5502 · outbound

This paper cites Recasting generic pretrained vision transformers as object-centric scene encoders for manipulation policies.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Recasting generic pretrained vision transformers as object-centric scene encoders for manipulation policies

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.462435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:714a9984a8ac9d4fe81fb07d2dd24f04d399b86d07bd03636bb913f19127f5fe

Observation f283615d-bf94-49ff-9622-d4e4bbf12879 · outbound

This paper cites Emerging properties in self-supervised vision transformers.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Emerging properties in self-supervised vision transformers

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.526227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:432a4d56bc465ddacbba563228bed9ad6516881deb12b811dc7e66ccdba0ee53

Observation 7d6bc277-e28e-48f5-b72b-4e0f35ddae03 · outbound

This paper cites Propainter: Improving propagation and transformer for video inpainting.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Propainter: Improving propagation and transformer for video inpainting

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.535049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:b3e468ddb3c95d08878e439e5ad8dd7fb00489d1e5c3d709cc9ab40cc946a3e7

Observation b5284360-4acd-42ed-ae5c-d73e26a84faf · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:11:03.882518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8ea1227b277d367863a247f10fb342ec07797209920d141592ca6030c816e5b8

Observation 9d340648-6b4d-4af0-b528-e43b0d83f2e9 · outbound

This paper cites Multi-view hand reconstruction with a point- embedded transformer.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Multi-view hand reconstruction with a point- embedded transformer

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.503815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:5d810325f8d5d22d799bc8eb033e38406bafeea878f89686f85aff9a7a6dbf2a

Pith citing papers

Observation 67af477c-5b1b-4692-b6b5-3501432580d5 · inbound

EgoRecovery: Acquiring Failure Recovery Ability Through Human Recovery Demonstration cites this paper.

EgoRecovery: Acquiring Failure Recovery Ability Through Human Recovery Demonstration LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T11:54:20.104031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:54:20.104031Z digest=sha256:31859f30bcfaff601b29b4e89222db0bc076dacb4166188a3b0bc327b443ea52