Pith. sign in

Paper Citation Record · LEDGER

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone

As of 19 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2607.25895.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25895 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T01:15:08.021366Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved67
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 147f4e00-5f78-40f2-89ff-928413c31031 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone RT-1: Robotics transformer for real-world control at scale

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.843794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.843794Z digest=sha256:03c622032c746afc6130aac16476ef2a4ad81159c626c547f8670e4af7b4e642

Observation 2e53a6d7-d331-4f09-b78f-87b7c1bca915 · outbound

This paper cites BridgeData V2: A Dataset for Robot Learning at Scale.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone BridgeData V2: A Dataset for Robot Learning at Scale

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.847028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.847028Z digest=sha256:856afb7ae806b10fd779a4984e528a0ea710b2d51f9b54d1f7fa8a1a3f730fb4

Observation ff3adf4c-4839-42e8-9da1-d0a358e67b63 · outbound

This paper cites DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.850750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.850750Z digest=sha256:9457487b40799a0d769bdcba79574d0d7ceb16c97ce2db237f25031b8d56ee20

Observation 868d0f32-df5c-46bc-a100-c44df49e1443 · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.853360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.853360Z digest=sha256:2a1b77e2c99e0ac03ed4f19ad44826dfdaa490086aab631bb3a7fd78153397d9

Observation 2e669a11-8af6-47c7-be3c-66313ca7aee6 · outbound

This paper cites RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.856378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.856378Z digest=sha256:87c43484cb13c257a09efd97951efb98919601a3e1acba389804c537b5d62025

Observation 6a673843-423f-47d1-8ca4-7b4c690c5fa3 · outbound

This paper cites Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.859324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.859324Z digest=sha256:6a9a93b0fa4619eed5b0d8357914a1bfc4d1d4bab4c946e2c36c63e046fbbc8b

Observation 0065c2d7-f858-452b-9cdc-f6af0a9b9a7b · outbound

This paper cites FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.861801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.861801Z digest=sha256:c7d505da9b2f8f5c351327708a4b0cf7ba0d88f126576a1a11bcf87f255ad80a

Observation f5273232-f96a-4635-a2e9-1e2fb399d211 · outbound

This paper cites Data Scaling Laws in Imitation Learning for Robotic Manipulation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Data Scaling Laws in Imitation Learning for Robotic Manipulation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.863974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.863974Z digest=sha256:12ea669de630459f94aed431722d687479a98aabdd305f8c0534843706825afd

Observation defe2762-7122-4579-a349-712df11df18b · outbound

This paper cites AirExo-2: Scaling up Generalizable Robotic Imitation Learning with Low-Cost Exoskeletons.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone AirExo-2: Scaling up Generalizable Robotic Imitation Learning with Low-Cost Exoskeletons

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.866868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.866868Z digest=sha256:59f5dca1f88030374b09e59436e91b911b20c2b010fc752a99a4ac3f1566a770

Observation 38b77d06-c153-4515-89e5-cbf42d0088f1 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.870173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.870173Z digest=sha256:b0c825e229e8e5dd46a0be571902dedebb555a533205b91f25638560173de6c3

Observation 1ee8d4e8-cd52-4cc4-9390-7f454d0ecc91 · outbound

This paper cites H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.873690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.873690Z digest=sha256:459e1258ee2421b46fce65a1289a3cc3c9b2215baccc27626c5109ba3c3d5252

Observation dcc07222-1174-4731-97f2-5d0ee71b9681 · outbound

This paper cites John, et al.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone John, et al

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.876478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.876478Z digest=sha256:be594db1c3200698a270c819576291bfff6737c81286a1db32294e4be5c4805e

Observation a30d030d-8c85-4b7b-9df7-40e800deffa6 · outbound

This paper cites XRZero-G0: Pushing the Frontier of Dexterous Robotic Manipulation with Interfaces, Quality and Ratios.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone XRZero-G0: Pushing the Frontier of Dexterous Robotic Manipulation with Interfaces, Quality and Ratios

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.878525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.878525Z digest=sha256:39d61940542e7f4c4f194aa8f7a2af601e2d784341e5b43d9c7599f435987e5b

Observation d5d3449a-cd41-4454-a54f-7c5a59611904 · outbound

This paper cites RDT2: Exploring the scaling limit of UMI data towards zero-shot cross-embodiment generalization.arXiv preprint arXiv:2602.03310, 2026.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone RDT2: Exploring the scaling limit of UMI data towards zero-shot cross-embodiment generalization.arXiv preprint arXiv:2602.03310, 2026

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.881171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.881171Z digest=sha256:bfd482e3616270e8d201983e7026b95fc4199ed7d8b9479ead0e3337d9533e75

Observation a424e8c8-f4dd-450d-b2af-3278cbff3aa0 · outbound

This paper cites RH20T: A Comprehensive Robotic Dataset for Learning Diverse Skills in One-Shot.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone RH20T: A Comprehensive Robotic Dataset for Learning Diverse Skills in One-Shot

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.883647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.883647Z digest=sha256:51c0ff13260b380e95320eb0425e82d5823f9238b6fae3b7a48fe04c801bdabe

Observation 978c0a20-2697-4f79-b889-5d8d981413a3 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.886373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.886373Z digest=sha256:791de812f5f2936f7fd2151002b49df91837beffebcab4566b7eb77914180452

Observation 9aa15162-6e2c-4724-a890-8a2645afcbad · outbound

This paper cites RoboMIND 2.0: A multimodal, bimanual mobile manipulation dataset for generalizable embodied intelligence.arXiv preprint arXiv:2512.24653, 2025.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone RoboMIND 2.0: A multimodal, bimanual mobile manipulation dataset for generalizable embodied intelligence.arXiv preprint arXiv:2512.24653, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.888913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.888913Z digest=sha256:8786a99d0b9ab9364f35400b0697cab0206c823ac8f70c7141eff8ad1f8ef9d7

Observation d44abac6-5073-4d1a-961e-d191ede4d4ef · outbound

This paper cites Ego4D: Around the World in 3,000 Hours of Egocentric Video.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Ego4D: Around the World in 3,000 Hours of Egocentric Video

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.890816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.890816Z digest=sha256:d5c914743073a00de36f66cf36f6fab48b72145acac02b25ef43da3b9172ec6b

Observation d0dddefd-ac1e-4f22-8201-0c093845a150 · outbound

This paper cites Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.893849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.893849Z digest=sha256:c7556b8f8bd9b5d01838fb45191f4e3c3fc3761d46102fb65c47f74c48728e97

Observation 7c36d10b-e489-4729-82f4-79904937a768 · outbound

This paper cites EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.897484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.897484Z digest=sha256:7d92955e75943a7c275d7679e6549a1d514a52d2988b065d2f1efd31518da755

Observation 9413fe39-8258-4e23-914b-831abc79836c · outbound

This paper cites FastUMI-100K: Advancing data-driven robotic manipulation with a large-scale UMI-style dataset.arXiv preprint arXiv:2510.08022, 2025.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone FastUMI-100K: Advancing data-driven robotic manipulation with a large-scale UMI-style dataset.arXiv preprint arXiv:2510.08022, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.900559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.900559Z digest=sha256:17442f7211a612a86f81f7aaf1cd1cc00dc76e04375d990627213d4997247580

Observation bf8f3ca5-a8cd-4e1e-8ba6-d69d3fdf730a · outbound

This paper cites VISTA: Vision-Grounded and Physics-Validated Adaptation of UMI data for VLA Training.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone VISTA: Vision-Grounded and Physics-Validated Adaptation of UMI data for VLA Training

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.902567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.902567Z digest=sha256:8e9614a083521ce478027354122e94f779b669ec8fa7622b451dde968e14f78c

Observation b8585936-bad7-4706-8088-9af61977cd07 · outbound

This paper cites DexCap: Scalable and Portable Mocap Data Collection System for Dexterous Manipulation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone DexCap: Scalable and Portable Mocap Data Collection System for Dexterous Manipulation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.904750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.904750Z digest=sha256:9f5b7780c72d87f143fdccd8c4601230749d9b5c0c752aad3aefa2021a73927b

Observation d193376a-1e00-4fe0-93b3-545ef6bb2a9e · outbound

This paper cites DexUMI: Using human hand as the universal manipulation interface for dexterous manipulation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone DexUMI: Using human hand as the universal manipulation interface for dexterous manipulation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.906926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.906926Z digest=sha256:aa5f4d61ed40c8f9e849c95dfa81f802b9e220834320e798f1759a14631bf2d3

Observation 96b24f2b-8238-4b2a-949e-addd6878a1cb · outbound

This paper cites DexWild: Dexterous human interactions for in-the-wild robot policies.Robotics: Science and Systems (RSS), 2025.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone DexWild: Dexterous human interactions for in-the-wild robot policies.Robotics: Science and Systems (RSS), 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.909567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.909567Z digest=sha256:99d60d6e02f95d11609e88f2513c4f5f20548d3ae841b37b061990f47eb78089

Observation a65d6843-15a8-4548-9b56-3878c77c3b24 · outbound

This paper cites ARCap: Collecting High-quality Human Demonstrations for Robot Learning with Augmented Reality Feedback.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone ARCap: Collecting High-quality Human Demonstrations for Robot Learning with Augmented Reality Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.911830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.911830Z digest=sha256:f918aec0e2ff4902085f9ae4525fa15f5426dce47ecf6a790dc50c3689099696

Observation 49a4a3bb-e427-46a3-8ce8-4eae216f2871 · outbound

This paper cites AirExo: Low-Cost Exoskeletons for Learning Whole-Arm Manipulation in the Wild.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone AirExo: Low-Cost Exoskeletons for Learning Whole-Arm Manipulation in the Wild

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.914989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.914989Z digest=sha256:b5be99bc0f195107703683c27d60097eb0a6b2fd8375356a5973ad9a4aff8bf0

Observation 5815b737-2195-4193-96c6-cae032e081e4 · outbound

This paper cites EgoMimic: Scaling Imitation Learning via Egocentric Video.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone EgoMimic: Scaling Imitation Learning via Egocentric Video

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.917085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.917085Z digest=sha256:c2139cc6c22f2a514aa68131a54e633f229b596afab4182fd36e4c452aec7c8c

Observation 30ddacf8-d7ee-4b88-879c-c965b75954f4 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.920021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.920021Z digest=sha256:bdac83098d261a96ea70aee5a20acd584135d004bae042f986fe588de04e45d9

Observation 38f3699b-b102-45f2-b4cd-aeb40d907186 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone OpenVLA: An Open-Source Vision-Language-Action Model

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.922128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.922128Z digest=sha256:bc520c88ce4c51bd1dab2a45ed4832ca0884c16566ad135b6547cfbf9ec81be1

Observation d9baf7a6-d52e-4c69-817e-dd3f3be8db15 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Octo: An Open-Source Generalist Robot Policy

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.924123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.924123Z digest=sha256:89a7551eb3c11c8a33f29dd30ad57b275a38cd76beabcd1c7204bb1c45af552f

Observation e74a2ee7-6405-4057-9062-26a35cb849ec · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.926972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.926972Z digest=sha256:0e218e6063cd373c89cd9bc02c9077d09969bfbd220cfa261e545b8008716012

Observation 67144df1-0593-4214-b1ea-cc967f0a7428 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.930670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.930670Z digest=sha256:fdbecea1d2427c58565d141deb54837317417c1dd418b19960de46e859722dfd

Observation 5ae5ab42-a1ff-483c-9906-505517c03b38 · outbound

This paper cites GR-3 Technical Report.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone GR-3 Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.933554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.933554Z digest=sha256:ffd74eac662f5cb8c4e6e59ed34c6e5e3246d5f30c66de0bfae176ffa2d6ca28

Observation 56b39639-8319-445a-a65c-14ad0020a69c · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Gemini Robotics: Bringing AI into the Physical World

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.936121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.936121Z digest=sha256:d67d31e9c629dbfe3496081c9a3f70070fc29c3d7e437a4e8d978fd919712bee

Observation 088e8c96-af70-488b-a483-af6c0433a75e · outbound

This paper cites SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.939059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.939059Z digest=sha256:fbd404ef029ba4507896bcb254cb6351e4d096840138695c01f4c47c31328985

Observation ed28d275-4394-45f9-acf0-12b20f40b23e · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.941316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.941316Z digest=sha256:a019382fad38cb15aa9f174aef85f4f6439f7ceff5816d4b51571bee7b65afde

Observation d7ad9593-b41a-4079-a6ee-15b382734c49 · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.943496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.943496Z digest=sha256:4057fe97712e959a0e998f26d56f68294c7b3a375c3c0f27705014f79b842a8f

Observation 59b63377-e4f4-4349-a74c-a63399492d06 · outbound

This paper cites A Pragmatic VLA Foundation Model.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone A Pragmatic VLA Foundation Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.945870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.945870Z digest=sha256:9f1f8864edfbb80962a3010c83470a413e0c66a120acd0081650a7cb801b4842

Observation 25f985fe-dfa6-4f5e-a05d-30e2e6cbfdda · outbound

This paper cites Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.948972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.948972Z digest=sha256:ce7bfba72b1da8ba849d359da9ba69967603e4f3b1080de566713ca08810e760

Observation 6b6e4223-43cf-4315-afdc-a7dad6cc1d00 · outbound

This paper cites Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.952361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.952361Z digest=sha256:511980cdd0f0bc84cf4ad6a9d6b838b350fed2024285b5b97d4b943e832dfa65

Observation 4adee39c-a9a6-4fec-b7e9-f6ac4c13d4e5 · outbound

This paper cites Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.955936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.955936Z digest=sha256:463fd8e0f2f027fa55b501c6283b44b6e31cc56f24a3d1dc0fae77e8837c6bf0

Observation e25a7d0b-8ea0-4e5f-aa72-88f322eec3c9 · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.959147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.959147Z digest=sha256:0f42de0069097a83d4cadee31c7c22535e1e5b2d44519a45d07463024086ad7d

Observation b9d5f229-40a8-4cdc-bfc8-0a0827f763b4 · outbound

This paper cites Learning Universal Policies via Text-Guided Video Generation.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Learning Universal Policies via Text-Guided Video Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.962162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.962162Z digest=sha256:527cbb4597e2f62cc964dc471801267b14ee09fdbd6f90f639a9c02a0cd92730

Observation 516c59fd-800a-454f-99fc-d29546754d0c · outbound

This paper cites Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.964827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.964827Z digest=sha256:ad20ef0b3a2fee561dec80e0d86d3dd78907c9d358c3a7026a7c533775b42cff

Observation 1d799c39-a199-4718-ad51-e6ac0467de3d · outbound

This paper cites World Action Models are Zero-shot Policies.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone World Action Models are Zero-shot Policies

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.967585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.967585Z digest=sha256:c3735a6750116cfc0f2242f45ab7488923710af551ee566fab8711f6f885c8cd

Observation 2074a2cd-9842-45f5-aee4-a4097fe86891 · outbound

This paper cites WorldVLA: Towards Autoregressive Action World Model.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone WorldVLA: Towards Autoregressive Action World Model

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.970203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.970203Z digest=sha256:c3c4925b678b889ffc107127b83c07d8c2e1973d181795f5b7124957f6e863a3

Observation ce9ae302-c28f-4fed-9c54-524a75b79db7 · outbound

This paper cites Causal World Modeling for Robot Control.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Causal World Modeling for Robot Control

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.972214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.972214Z digest=sha256:60d5c61cdf88ac1e5c52cf318adb86da7ce4c008406676a93e8959d3efba51b2

Observation 591bb4ff-12c0-45ac-a384-9244297150fb · outbound

This paper cites StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.974934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.974934Z digest=sha256:90ef8c12ce18af35802c03a38e07122c0f4e6c5ab9496af58ab759a363091543

Observation b1a134b5-f62d-4f1b-87be-fdca0542261d · outbound

This paper cites VINS-Mono: A robust and versatile monocular visual-inertial state estimator.IEEE Transactions on Robotics, 34(4):1004–1020, 2018.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone VINS-Mono: A robust and versatile monocular visual-inertial state estimator.IEEE Transactions on Robotics, 34(4):1004–1020, 2018

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.978401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.978401Z digest=sha256:598044fd34148ea549dcb5e918d3233a2dc1c483da255f71672bbe1316a05a3e

Observation bad92038-1fbc-4068-b969-0084137e6691 · outbound

This paper cites ORB-SLAM3: An Accurate Open-Source Library for Visual, Visual-Inertial and Multi-Map SLAM.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone ORB-SLAM3: An Accurate Open-Source Library for Visual, Visual-Inertial and Multi-Map SLAM

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.980726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.980726Z digest=sha256:029f1da26144c352e335adaaa90d5eadd03eda06e49ef1584033ace7aeb3cc25

Observation 978a7d91-8f86-4280-9bec-a485e9d31743 · outbound

This paper cites AprilTag: A robust and flexible visual fiducial system.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone AprilTag: A robust and flexible visual fiducial system

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.982863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.982863Z digest=sha256:6dad964df47700ce6fc483f975a1b634474a73c555b5a6b5db42b1c766e889dc

Observation c96a28f6-be8b-4d46-a199-b9c6abd122d2 · outbound

This paper cites VersaVIS: An Open Versatile Multi-Camera Visual-Inertial Sensor Suite.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone VersaVIS: An Open Versatile Multi-Camera Visual-Inertial Sensor Suite

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.985215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.985215Z digest=sha256:31cd810419c80223442b0eab254e6b372a24cf199251745970c1b1a066511546

Observation bb3d007a-cdcd-497f-839c-527d7b23d054 · outbound

This paper cites DAS Fingers: A high-precision multimodal data-collection device for embodied AI.https: //www.genrobot.ai/products/finger, 2025.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone DAS Fingers: A high-precision multimodal data-collection device for embodied AI.https: //www.genrobot.ai/products/finger, 2025

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.988427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.988427Z digest=sha256:f1e314cfa3aa6bbcc300d2a5a0f0a8df2bea66c4cb6c2a06161b0604d9b7748c

Observation 69bf5fc2-54fc-4e38-b605-466084a04a76 · outbound

This paper cites TacUMI: A multi-modal universal manipulation interface for contact-rich tasks.arXiv preprint arXiv:2601.14550, 2026.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone TacUMI: A multi-modal universal manipulation interface for contact-rich tasks.arXiv preprint arXiv:2601.14550, 2026

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.990542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.990542Z digest=sha256:21c83ff212c5ee24f060ee6f42e04ffb8e2824809268d5df3bea55d1a259a10a

Observation 236f8bf4-e2f2-4ba5-b28f-559660614910 · outbound

This paper cites DynaSLAM: Tracking, Mapping and Inpainting in Dynamic Scenes.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone DynaSLAM: Tracking, Mapping and Inpainting in Dynamic Scenes

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.992313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.992313Z digest=sha256:588a9404ab5719d7bfde91e6d3fc7b00c370b94129c96a77f83843d85f42085b

Observation 275e6924-0abc-4817-9e5c-89193e7a18b6 · outbound

This paper cites UMI on Legs: Making Manipulation Policies Mobile with Manipulation-Centric Whole-body Controllers.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone UMI on Legs: Making Manipulation Policies Mobile with Manipulation-Centric Whole-body Controllers

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.995654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.995654Z digest=sha256:94f831ab10bafa9d14fc06f790311bcc60d7d6a7c2222abf8944cb0527c9f37f

Observation 0fa8150c-ae49-4866-aade-f4a9675ddd5b · outbound

This paper cites RoboVQA: Multimodal Long-Horizon Reasoning for Robotics.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone RoboVQA: Multimodal Long-Horizon Reasoning for Robotics

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.999020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:07.999020Z digest=sha256:3175d61abd3928f41194a3b203672e49687376563e6f7a1e8ea11d987b7eec5a

Observation 0a6cd157-ec62-4e7e-8ecb-e1199db56ee6 · outbound

This paper cites On the Continuity of Rotation Representations in Neural Networks.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone On the Continuity of Rotation Representations in Neural Networks

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.001712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.001712Z digest=sha256:d412d2ef8c8bf2b871ca8f8051abb4157b499a4f2b4e97662e6b39551385fad4

Observation 019fd61b-3910-4cf0-bd06-0edade591ee3 · outbound

This paper cites Qwen3 Technical Report.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Qwen3 Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.003800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.003800Z digest=sha256:a5ff084950f3e0df44180df562b1daf08458ebd836f1a1144278b38c0f948760

Observation 2bb3b13f-53a2-4b4f-9d87-c30234145da6 · outbound

This paper cites Flow Matching for Generative Modeling.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Flow Matching for Generative Modeling

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.006939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.006939Z digest=sha256:28ea7ed9e74993f9ffdfb3dd0a61712066f4249918ab7ae7f48254f1b2dd16fc

Observation db7e7c52-4076-4a3f-b2be-538915c83d31 · outbound

This paper cites Scalable Diffusion Models with Transformers.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Scalable Diffusion Models with Transformers

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.009966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.009966Z digest=sha256:9d748ae38086ba69782fa9392043850d123d57e09674acf745bbc5a6ea70f469

Observation 30ae9065-0513-4b23-a167-08b6ce834775 · outbound

This paper cites Zhao, Vikash Kumar, Sergey Levine, et al.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Zhao, Vikash Kumar, Sergey Levine, et al

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.012207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.012207Z digest=sha256:ec3bc2e6a19812abe31b3792b18c459ef0ba4bf17a9758a2857dd79c6a526cf2

Observation 641e5e1c-f00f-493e-91a4-37ace503c0a7 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Diffusion policy: Visuomotor policy learning via action diffusion

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.014056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.014056Z digest=sha256:b28bca7a3eb5a090ec56b891c7121480b5f0f0cbbc75b504c5e93d37f7f17afc

Observation 97b636d6-5d0e-4eda-b94e-1f510350e181 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone PaliGemma: A versatile 3B VLM for transfer

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.016418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.016418Z digest=sha256:f50b4fe0d28ce5135b870bbe39bd6413472a2eb0fa99a6f6f007dccca733d583

Observation 5bd9ee1d-28fa-4af5-aac0-a1fe9e9839f4 · outbound

This paper cites Decoupled Weight Decay Regularization.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Decoupled Weight Decay Regularization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.019104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.019104Z digest=sha256:475fd03f235dff051757749fb53286043251c8d8e281e0354f9219f85f746aef

Observation 146e38cd-4f6b-4456-8ee0-bef8b44e30e5 · outbound

This paper cites A careful examination of large behavior models for multitask dexterous manipulation.Science Robotics, 11(113):eaea6201, 2026.

HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone A careful examination of large behavior models for multitask dexterous manipulation.Science Robotics, 11(113):eaea6201, 2026

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:08.021366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:15:08.021366Z digest=sha256:8b112919ba2e18916511fa54797364864f140092efec5ee98f58188b080b7424

Pith citing papers

No inbound Pith citation observations are available.