Pith. sign in

Paper Citation Record · LEDGER

AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:1912.04443.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1912.04443 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T10:39:38.257200Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T02:26:26.937662Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1bf7b9b3-80c1-40ef-b636-f8128555e79f · inbound

Open X-Embodiment: Robotic Learning Datasets and RT-X Models cites this paper.

Open X-Embodiment: Robotic Learning Datasets and RT-X Models AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:23:24.368656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T17:23:24.255829Z digest=sha256:34524bf4a15e42e033afbb2f495c0c19bbfa5d36fd2d7f2797d31b7db8a0381d

Observation ce900b45-d18b-473b-b51d-ed8575245c92 · inbound

Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation cites this paper.

Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:02:55.408431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T22:02:55.240949Z digest=sha256:636ae2f892255a73f8c0e90163b25c342569e5f950f31adc92ac0dddd27f146b

Observation f19d7843-b3b5-4607-adb3-4e6d1a10c0af · inbound

DIRIGENt: End-To-End Robotic Imitation of Human Demonstrations Based on a Diffusion Model cites this paper.

DIRIGENt: End-To-End Robotic Imitation of Human Demonstrations Based on a Diffusion Model AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T10:39:38.257200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T10:39:38.257200Z digest=sha256:2cea3cbf18fbed2c719cf4b1de28a384246a2c96168dc25e5f435a1d3e851976

Observation 52287843-3473-4c86-9fc4-9eb2057dcd50 · inbound

Preference VLM: Leveraging VLMs for Scalable Preference-Based Reinforcement Learning cites this paper.

Preference VLM: Leveraging VLMs for Scalable Preference-Based Reinforcement Learning AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-09T14:52:27.279746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:52:27.279746Z digest=sha256:e4eab2ed978a8c43fe0c8b09f3e82f5d9d2464055883855afdc79826156813f4

Observation 2d337009-913c-4a74-a656-5e40905d45cd · inbound

Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt cites this paper.

Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:26.260839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:51:26.260839Z digest=sha256:8db87d6154439fe954673d271ab58dc7be93a2a5fc088c322b1acdcab493f20a

Observation f69a72bb-c9be-40b5-a4f1-d5bc339b792b · inbound

Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations cites this paper.

Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 107

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:37:07.499088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T06:36:13.144868Z digest=sha256:7726a75aa92f024c1eaa53a44c7fbf31c51d81820604400e9f78d5009fad8686

Observation cfd2d52e-fea6-4740-9ea9-f480858a177f · inbound

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data cites this paper.

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:03:01.005282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T17:02:18.358675Z digest=sha256:0d0064a85a7b48157ae74961ecbbd05d467776050cb5323902910fb4c41ac97f

Observation 12263698-48a1-4442-89f7-a4484a809f74 · inbound

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World cites this paper.

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:35:56.985553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:07:41.489995Z digest=sha256:585228956eaa4731042217bff1dae151e65a788e733d4bfd4aef0b2d36e2b4fb

Observation de5c3b6d-4c81-4b61-83b1-916bb357f3a3 · inbound

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World cites this paper.

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-13T08:25:22.011013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T08:25:22.011013Z digest=sha256:64ae7c06c9be545992eacd69a050607c13678ccd449468e7eaaf57b0a6f3f66f

Observation 6abbc5bd-eaa8-43ee-b1e2-31fd4b2781e8 · inbound

Reinforcement Learning from Cross-domain Videos with Video Prediction Model cites this paper.

Reinforcement Learning from Cross-domain Videos with Video Prediction Model AVID: Learning Multi-Stage Tasks via Pixel-Level Translation of Human Videos

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:26:26.939589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T10:54:49.582908Z digest=sha256:bf8027bac94ddac964999cc384e0bb8ab986e1bb6fd19fda1907f0f22fd467c6