Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T04:32:58.733165Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 83 of 83 outbound references and 72 inbound Pith citation observations for arXiv:2507.12440.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T04:32:58.733165Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:20:20.579410Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
83 of 83 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
Observation 0ed6bc88-415e-4e99-aafd-20845899eef2 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Vuong, S
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9c44f7f5-a543-4b4c-a5a6-bddec513fd05 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Khazatsky, K
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c93141f6-9c67-4065-9b40-fefb354c28be · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos OPEN TEACH: A Versatile Teleoperation System for Robotic Manipulation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8b923843-f2b7-45a8-87c7-0eba72127999 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos TeleMoMa: A Modular and Versatile Teleoperation System for Mobile Manipulation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a640a85a-966f-49b3-b0a0-79ee35a61d48 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation aab5372b-9d8d-49ba-a048-16f4e16b30bf · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7b4ddec7-c528-4651-8ff7-608598c44a8a · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Fang, H.-S
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ada99f2f-06e6-4937-9026-67a493d30ded · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1bae9bd2-4dd4-49de-997e-52e812636c92 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Naceri, D
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9032e7d9-70bc-48af-be0e-d67cbcff8383 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Open-TeleVision: Teleoperation with Immersive Active Visual Feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 23390a69-66f6-41a6-931b-fa9e32351cba · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Bunny-VisionPro: Real-Time Bimanual Dexterous Teleoperation for Imitation Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d32eb02a-43ed-4527-b766-9bf74f0ff7f4 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Gaze-guided hand-object interaction synthesis: Benchmark and method
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0ee0858d-8750-4c2c-85e8-21a43172222c · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Ghosh, H
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d76022e1-97c2-4756-b9e5-95ff04cc5f06 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e5db4faa-709d-4796-8112-becec2c57d6a · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Black, N
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1573cea1-701d-47a8-ac5d-8790141f15b4 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Brohan, N
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3b146c72-5b74-48c8-bed7-50e10250e9fd · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Isaac sim: Advanced simulation for robotics development
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bdcdb99b-d1c1-497a-86e9-fd906deb400b · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Romero, D
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9539cc4c-6c4c-4019-a970-da5e8b7ff12b · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Rodriguez, M
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7dfe254e-eabd-4e75-a992-e1a46506b1f6 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Rosales, R
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f871a194-e135-4782-8743-cc76b36d39f5 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Prattichizzo, M
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2b979ad0-6171-4519-abce-ddcb2843364b · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Ponce, S
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a79c1865-d90a-44e4-ac4a-9e723cbd5e2c · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Ponce, S
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 811a25d3-444d-4074-8465-009786708e4f · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Zheng and C.-M
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e7ad1e08-c8bf-49f2-8c5e-ae013f93e608 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 038af182-6a54-4ba4-a5e9-e0af5dbf2429 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1238de21-5c5e-4e1e-86db-2dd1c7005122 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Nagabandi, K
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8a6a568e-ff5e-48ec-8e50-196e0d7136ca · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Jiang, S
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 36b87ba5-ec61-4175-a493-59555a067100 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Corona, A
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e57fd03-f052-41d9-a27e-db9a4e676531 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 613eea68-9074-4920-a4be-7060e7de6a3c · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8f0b2e07-4e1b-43aa-bdd2-04947160a258 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 98748107-6853-49e5-b002-d5e8b398b198 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Brahmbhatt, A
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e2bc1f8-2562-49bf-bf0f-53cc8645c3e2 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Turpin, L
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 35db358e-6c79-48ff-9af3-c40788dbeff2 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0baa12ed-e81e-4785-9fb6-4e04cce48a93 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Yoon, Ryan Hoque, Lars Paulsen, Ge Yang, Jian Zhang, Sha Yi, Guanya Shi, and Xiaolong Wang
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7f6ecd0f-6c98-428b-a03e-ef603ccb88c4 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos EgoMimic: Scaling Imitation Learning via Egocentric Video
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b31a7f52-4eb5-4ff9-bab5-1545dae98eff · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Achiam, S
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 865a6579-9aae-404e-961d-abe09e76db61 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 16ae8019-65d2-430c-9454-ff145993fb34 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9b81d31d-b7ad-44c8-abe3-376da3b75bc4 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Pratt, I
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e4eaae4f-6425-4526-9d66-077d003cf108 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Alaluf, E
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eab2daca-4438-4827-afb3-4ff0a3abaca6 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b3bab134-7c4d-4926-918e-e615bab6e6e1 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Huang, S
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7e40dc2c-45d6-4952-be4a-661bde1a0830 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 25767cf3-3856-497e-9de7-3690fe020f14 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c250a78b-7295-4aad-a829-6045709cac23 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7dba06be-5bb9-4ac7-8204-ee09a4721e96 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Mandlekar, Y
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d125605d-8cfc-48de-a49d-ff9bff0b9ac2 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Mandlekar, S
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 00f86eea-1d94-48df-bc8a-7c31ac50b801 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Dasari, F
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fb0f664c-479b-41ba-a2e6-643f98d96d53 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Kalashnikov, A
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c9a26a40-af76-4340-a70a-fc35b26658f6 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Damen, H
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ef81e9b-4bf1-4f77-9598-24af79387c88 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b0bba765-2663-4f9a-8912-f94e8ad932a4 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c2894ad7-535c-48ec-8c1d-ada1371d6031 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Grauman, A
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 276e07fc-29d1-405e-b6fe-7d69ceadd9fe · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Grauman, A
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4eaf3de1-56dd-4111-82f6-250aef88e7a4 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Mahdisoltani, G
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3e26c793-cd4e-4c95-9fa6-29d38e2a2e54 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Damen, H
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bde0bcb1-1ce9-4257-99b0-554f30c9b210 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 709e3014-4166-4e07-9d2d-b58736559dc5 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f91308d9-9a25-4f82-9dfd-7b17254ed7ee · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos R3M: A Universal Visual Representation for Robot Manipulation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f2633556-5944-414a-97cd-90aa1509c3bd · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Majumdar, K
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 17691082-9687-4709-a809-a9494f6760e0 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Karamcheti, S
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6b35b50a-abcc-4d3d-8c11-6daeeeac83e6 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Spatiotemporal Predictive Pre-training for Robotic Motor Control
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dd9b74d3-11c6-42be-86e6-da0f5e495af5 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Learning Manipulation by Predicting Interaction
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 162c26cd-a0a7-4313-83cb-fec5c051812f · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Latent Action Pretraining from Videos
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a3e33c1f-b5b8-42cd-b82c-98ab774ed1f7 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Lirui, C
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e3b33f33-822a-4f75-b31c-5aaa08be4c8f · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos NVILA: Efficient Frontier Visual Language Models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9ffaaa22-2da9-45f5-a5df-aee42c07c9b0 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos On the Continuity of Rotation Representations in Neural Networks
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7ac3faed-88bc-4171-87c8-d9726ec89d29 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Evaluating Real-World Robot Manipulation Policies in Simulation
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3e4ffae3-17a8-4163-8553-271383b92846 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Orbit: A unified simulation framework for interactive robot learning environments.IEEE Robotics and Au- tomation Letters, 8(6):3740–3747, June 2023
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 435674a3-24b8-4773-843f-ad7925994291 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e6f1e820-eb4b-4566-bfdf-602ac649fd5a · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Robotics
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7e58ed22-46f9-4e21-b05f-e05154f06b29 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 658dc08a-a6d9-4897-b482-b56c658884d7 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 86371c3f-7f10-4e18-ba62-39d56061c571 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3f450265-fcd6-4a08-b940-2dd66dbf4a27 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6bd75f0f-7626-47c2-880c-6b59314dc797 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos TACO: Benchmarking Generalizable Bimanual Tool-ACtion-Object Understanding
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e0d8c426-7f7d-4c62-8b9b-a8c896f144c2 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 30c213a6-8018-4c72-ae15-738af907cb88 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Oquab, T
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d737640d-a9a3-4478-8500-6b90a9fd273f · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Darcet, M
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d2fc7b23-a0d4-44d6-b283-0acab4b130f8 · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos Nvidia omniverse: A platform for virtual collaboration and real-time simulation
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2e1729d2-b841-4302-83cc-4fd095e8512b · outbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos 22 14 Appendix 1 Dataset Details Language Label: The combined dataset includes ego-centric RGB visual observations, wrist poses, hand poses, and camera poses
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c8eb7e53-f544-443a-bdd0-da4c063eaf84 · inbound
HERMES: Human-to-Robot Embodied Learning from Multi-Source Motion Data for Mobile Dexterous Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78290916-5698-45fb-8855-3e5e8279fec2 · inbound
Dexplore: Scalable Neural Control for Dexterous Manipulation from Reference-Scoped Exploration EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f85aafc8-57c6-4c01-9bd7-9893a48d9ece · inbound
Scaling Cross-Embodiment World Models for Dexterous Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad14006c-7330-461a-b4c3-6483b931bc30 · inbound
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 111
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e6e77eb6-0f66-40f4-8c65-6b9edf9ecf2a · inbound
EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3442e1a-da2f-43de-8da1-61d77661401e · inbound
LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 932f7c20-2c5a-4502-a56b-5c850330bb7d · inbound
NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26322bfc-c4db-4a3b-833b-27c7c5b48b45 · inbound
HoMMI: Learning Whole-Body Mobile Manipulation from Human Demonstrations EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f1c6ebd1-fa47-467f-b286-b9573dd6dcfa · inbound
AnyHand: A Large-Scale Synthetic Dataset for RGB(-D) Hand Pose Estimation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff72e32-ae8a-4f7c-b80e-ed568a4dca57 · inbound
Grasp as You Dream: Imitating Functional Grasping from Generated Human Demonstrations EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 65dbfa02-8a97-4b70-bb3e-fe86bc17275d · inbound
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 565988b3-96fb-4146-9501-9d899833eeea · inbound
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c74a357c-428d-45f6-8605-4ff72320bcfb · inbound
HEX: Humanoid-Aligned Experts for Cross-Embodiment Whole-Body Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0f6b5d9a-0edc-478b-8ebd-8942f31965e1 · inbound
HEX: Humanoid-Aligned Experts for Cross-Embodiment Whole-Body Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e6a4e8d-c484-4181-a36a-80d8c798b6a9 · inbound
ActiveGlasses: Learning Manipulation with Active Vision from Ego-centric Human Demonstration EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2f849ef2-1c7a-4974-be45-c6cf61e388d2 · inbound
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5b3a441a-bf1a-4da0-905a-6e077ac80f7e · inbound
${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ab2236a3-384b-4c1e-a969-5621dad563a7 · inbound
UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 89107eb2-2940-4a2e-b0a6-e3085a931b34 · inbound
CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a9991ad8-1c8d-4bfb-9443-9dfe0f7e1350 · inbound
CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56074592-06cd-44b0-be07-313183bcd6e6 · inbound
GazeVLA: Learning Human Intention for Robotic Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 97742f1d-b1dc-4bc0-a5b0-282eab0d5f3c · inbound
EgoLive: A Large-Scale Egocentric Dataset from Real-World Human Tasks EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 33c0d231-2c86-4a9f-be89-b633977734a6 · inbound
Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 142db0f4-95a6-4790-bbda-d1ab0ad3fb2c · inbound
Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cd341fd6-ad13-4b0b-aa5d-476c670cb09c · inbound
Being-H0.7: A Latent World-Action Model from Egocentric Videos EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 40e74a6d-e5a1-44f1-940d-46776c42825a · inbound
OmniHumanoid: Streaming Cross-Embodiment Video Generation with Paired-Free Adaptation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c48085dd-ca5d-4c85-b219-fdb75ce68d75 · inbound
World Action Models: The Next Frontier in Embodied AI EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 204
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8d727f0a-fc92-4900-b9dd-1d7d4b590b40 · inbound
SCAR: Self-Supervised Continuous Action Representation Learning EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fa2b2816-8762-4c3c-9096-c4e43be05bc6 · inbound
EgoKit: Towards Unified Low-Cost Egocentric Data Collection with Heterogeneous Devices EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0f7901b9-33cd-4cae-b2a0-1a5fa9e5c865 · inbound
StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fabb0792-6596-47fa-b9dc-4b6e2b2624f8 · inbound
Dexora: Open-source VLA for High-DoF Bimanual Dexterity EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bb346e45-957f-43f9-ac86-1dfd7103aeb5 · inbound
Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f10432fc-2773-45f3-943a-838afbd94419 · inbound
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e6b7e8f7-4376-472e-abb6-ceb3f62b59b9 · inbound
Grounded 3D-Aware Spatial Vision-Language Modeling EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b95a5978-599b-4eb8-a88c-ceba7b7e504b · inbound
From Human Videos to Robot Manipulation: A Survey on Scalable Vision-Language-Action Learning with Human-Centric Data EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f79ce0d9-38b5-4985-ac81-51426ad7fb61 · inbound
ActiveMimic: Egocentric Video Pretraining with Active Perception EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a8d1bbf8-b4ed-49f9-8a3d-2034fd8e6545 · inbound
What Matters When Cotraining Robot Manipulation Policies on Everyday Human Videos? EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eb25e211-2b4f-40c8-9972-461e2ad3a8ee · inbound
SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d3656c35-33f6-4663-8b7d-c25e52ed22ef · inbound
EgoPriMo: Egocentric Motion Generation for Interactive Humanoid Control EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1a6ae559-f93f-4a7e-a9fe-577820a3819d · inbound
MotionWAM: Towards Foundation World Action Models for Real-Time Humanoid Loco-Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ccc259b1-1094-49f7-9d6e-f222ebc55d83 · inbound
LUCID: Learning Embodiment-Agnostic Intent Models from Unstructured Human Videos for Scalable Dexterous Robot Skill Acquisition EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation df504181-4c65-40e4-9a47-4e05af4bb914 · inbound
$\mu$VLA: On Recurrent Memory for Partially Observable Manipulation in VLA Models EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f1340118-19e1-4b7d-8edd-c71c044bbf4c · inbound
EgoEngine: From Egocentric Human Videos to High-Fidelity Dexterous Robot Demonstrations EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7922ea34-3b38-457b-8fbb-cdafda1e772f · inbound
EmbodiSteer: Steering Embodiment-Agnostic Visuomotor Policies with Joint-Space Guidance for Zero-Shot Cross-Embodiment Deployment EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3f93b938-4c8d-4bdf-860d-f854ff00101c · inbound
FTP-1: A Generalist Foundation Tactile Policy Across Tactile Sensors for Contact-Rich Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b0a28147-1a24-4c47-aa17-1de436a5d9a1 · inbound
Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a21ea5b0-d478-4cd8-bc39-fef362b3d092 · inbound
T-Rex: Tactile-Reactive Dexterous Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c8fa37e1-ecbb-49f1-ad94-e889a471146d · inbound
ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b2c97fd8-6c75-4771-9235-58d325786214 · inbound
ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3f728eff-8671-4dce-b4ad-0d0782c52100 · inbound
MolmoMotion: Forecasting Point Trajectories in 3D with Language Instruction EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fcb465fd-1fb9-4042-a1a4-aaf2326d7ce3 · inbound
Do as I Do: Dexterous Manipulation Data from Everyday Human Videos EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3c0ed274-eed9-4717-a7ff-a1651544ec89 · inbound
Robot Self-Improvement via Human-Video Dynamics Models EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0e10fb32-2be9-4c54-a351-15c785ad5163 · inbound
Wh0: Generative World Models as Scalable Sources of Egocentric Human Hand Manipulation Data EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3f022312-5323-40e6-839e-a5bdd55acc8f · inbound
OpenHLM: An Empirical Recipe for Whole-Body Humanoid Loco-Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a93b0869-f00c-4480-988d-94f1e06b26b1 · inbound
LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 363bcb00-4a44-42cd-923b-dd79a16e2d23 · inbound
PointVG-R: Internalizing Geometric Reasoning in MLLMs for Precise Pointing Localization via Visual Chain of Thought EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7431e20f-57f7-4d3e-accf-5bc403869779 · inbound
Toward Low-Latency Vision-Language Models with Doubly-Correct Predictions in Egocentric Visual Understanding EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b3f5f51c-379b-4bcb-82aa-967b6acabd28 · inbound
Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly? EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d83d8a1e-aa7f-475f-804c-21e5f767c959 · inbound
Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly? EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69eeb5ac-c36e-4839-b093-e724b23c8af2 · inbound
Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5fde026c-793c-4a22-9fe3-73798e4de340 · inbound
Human-as-Humanoid: Enabling Zero-Shot Humanoid Learning from Ego-Exo Human Videos with Human-Aligned Embodiments EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 12b12969-e47f-4eba-906d-ee8abef1d384 · inbound
EgoGapBench: Benchmarking Egocentric Action Selection in Multi-Agent Scenes EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5214ee80-de4c-4955-bb8e-b7b497330267 · inbound
Human-Centric Transferable Tactile Pre-Training for Dexterous Robotic Manipulation EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 18ce32f8-ccd3-4bbb-83fe-7fbd7bb17eda · inbound
WSA$_1$: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a02135a-b814-4c23-a705-10d9051c8fde · inbound
EgoWAM: World Action Models Beyond Pixels with In-the-Wild Egocentric Human Data EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f1521096-62c9-424b-8ad1-8040816255d5 · inbound
EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a58bd9c8-4c2a-4b09-9131-e0a0307058df · inbound
Open-AoE: An Open Egocentric Manipulation Dataset and Toolchain for Embodied Learning EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4c67866-0373-46da-8f5a-9db043203143 · inbound
MEMORA: Embodied Action Memory from Egocentric Videos for Reasoning and Planning EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45e4ddb7-eb0d-4b11-adcf-24fcbb078b2f · inbound
Exo2EgoPose: Leveraging Exocentric Demonstrations for Vision-Language guided Egocentric 3D Hand Pose Forecasting EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19195c38-fbae-483b-915e-b10e7145e626 · inbound
EgoRecovery: Acquiring Failure Recovery Ability Through Human Recovery Demonstration EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48498a43-221c-4383-9281-9d9f02c6db98 · inbound
Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0726349b-97cf-43ef-854b-8890bac421c7 · inbound
Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.