Pith. sign in

Paper Citation Record · LEDGER

Hand-Object Interaction Pretraining from Videos

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2409.08273.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.08273 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:33:47.325060Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:59:19.461553Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a39f1f6b-150b-463c-813d-f6feffc9c986 · inbound

DextrAH-RGB: Visuomotor Policies to Grasp Anything with Dexterous Hands cites this paper.

DextrAH-RGB: Visuomotor Policies to Grasp Anything with Dexterous Hands Hand-Object Interaction Pretraining from Videos

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T10:56:18.962304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:56:18.962304Z digest=sha256:b587b72cecd51b5c92aa9e9b271fda35e7537487377f9816a26f95c5b81dcd13

Observation 5ce49807-3e8e-46a9-87c7-5ab79d51082a · inbound

FAST: Efficient Action Tokenization for Vision-Language-Action Models cites this paper.

FAST: Efficient Action Tokenization for Vision-Language-Action Models Hand-Object Interaction Pretraining from Videos

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:52:32.070856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T08:52:31.686474Z digest=sha256:9a818b29f3ec5f8b0347e07343a77dc04733368210cf87571b02cf33bb770795

Observation 018c4148-5f1a-4888-8da3-04f02e8e60c9 · inbound

Crossing the Human-Robot Embodiment Gap with Sim-to-Real RL using One Human Demonstration cites this paper.

Crossing the Human-Robot Embodiment Gap with Sim-to-Real RL using One Human Demonstration Hand-Object Interaction Pretraining from Videos

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T12:33:47.325060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:33:47.325060Z digest=sha256:cb55d738d157eaf6a171b50917f3b5a431cc0ca0079843a8f86bb6f2d9c438aa

Observation a485077d-05a2-48aa-a7d3-6b0ff4bdcec6 · inbound

Visual Imitation Enables Contextual Humanoid Control cites this paper.

Visual Imitation Enables Contextual Humanoid Control Hand-Object Interaction Pretraining from Videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T23:48:57.786901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:48:57.786901Z digest=sha256:674bc70cd6fcb2eeca6298615fd389ee841f9ec386569e3b11edc019bdc6320e

Observation 1da573a4-fa16-406e-b0da-9fd85a847504 · inbound

Web2Grasp: Learning Functional Grasps from Web Images of Hand-Object Interactions cites this paper.

Web2Grasp: Learning Functional Grasps from Web Images of Hand-Object Interactions Hand-Object Interaction Pretraining from Videos

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:32:05.485477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:32:05.485477Z digest=sha256:cf6c43519f988611dbd11431ba87acc83894ab166a88c6653f6bc98508fb8b04

Observation b923f554-5c43-4fdb-b8db-d14060557541 · inbound

DexWild: Dexterous Human Interactions for In-the-Wild Robot Policies cites this paper.

DexWild: Dexterous Human Interactions for In-the-Wild Robot Policies Hand-Object Interaction Pretraining from Videos

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T15:21:44.605498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T15:21:21.778285Z digest=sha256:5921a5eaf17d7e0927fb0d6b9e73bcc6432e8dee9b74d5178af6e572a150ff56

Observation 84116076-050a-4480-9b81-b1812e658e19 · inbound

Emergent Active Perception and Dexterity of Simulated Humanoids from Visual Reinforcement Learning cites this paper.

Emergent Active Perception and Dexterity of Simulated Humanoids from Visual Reinforcement Learning Hand-Object Interaction Pretraining from Videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:51.592709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:51.592709Z digest=sha256:c9b2914d0078b18a8d5724bdc04d35743a2c5c0bde410b20c660f89d23cae174

Observation 48dea0a4-fb80-46e4-b1b0-f2698a74f44a · inbound

EgoZero: Robot Learning from Smart Glasses cites this paper.

EgoZero: Robot Learning from Smart Glasses Hand-Object Interaction Pretraining from Videos

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:01.013692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:01:01.013692Z digest=sha256:d85d8f02c27e1550e9f1dc3534731c9d5c8a388976bd5c04f4886120cd96e722

Observation 5ed1133c-a3bf-4f78-b182-ae50846e14a6 · inbound

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios cites this paper.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Hand-Object Interaction Pretraining from Videos

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.683935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.683935Z digest=sha256:14f81d9e2f9930a78892253ff09c5a731ff121274fd14e36a3fa13ebc3d4f57e

Observation a1a452af-a3dc-4e44-9114-645146ce12b3 · inbound

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos cites this paper.

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos Hand-Object Interaction Pretraining from Videos

Reference 131

Resolution
unresolved
no resolver link, observed 2026-08-06T15:33:48.852618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:33:48.852618Z digest=sha256:9b3e03930a8405bc0533b51782cc0a15794d5b12d9456f47616195a545fe30bc

Observation 0f290199-dda1-4901-8ee0-970da8e9dfc5 · inbound

Do as I Do: Dexterous Manipulation Data from Everyday Human Videos cites this paper.

Do as I Do: Dexterous Manipulation Data from Everyday Human Videos Hand-Object Interaction Pretraining from Videos

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:59:19.464136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T20:51:21.209882Z digest=sha256:24dabf9456dd4b35b19a9d2e5f5098015e82697e4142335b86ff558c0ecc7bcb

Observation d03231c7-7eba-4024-bf85-6157e199fc48 · inbound

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis cites this paper.

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis Hand-Object Interaction Pretraining from Videos

Reference 179

Resolution
unresolved
no resolver link, observed 2026-08-01T19:06:46.826275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T19:06:46.826275Z digest=sha256:c326d89c3135ca68dc288f0e1330e8932be6d1d07ceede1f1101a0784325297e