Pith. sign in

Paper Citation Record · LEDGER

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

As of 7 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 9 inbound Pith citation observations for arXiv:2509.09769.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09769 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T18:51:36.127656Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T06:48:14.554799Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T22:16:36.335996Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved59
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49b4c5c5-1134-42d7-890b-13f411560d03 · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.979246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.979246Z digest=sha256:4c9a451d12748b879124a87f7fd207430355b2f4352e03a7f8772c0af7a6e33a

Observation 8bcd6db7-8ff7-4cde-a333-c674cf4f8a72 · outbound

This paper cites One-Shot Imitation from Observing Humans via Domain-Adaptive Meta-Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos One-Shot Imitation from Observing Humans via Domain-Adaptive Meta-Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.982472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.982472Z digest=sha256:0c6e87c29cb05e6e14dc6632571bd60f444e41359f7ecbaa19130611016ff44a

Observation f5075199-53e9-4f50-bbbc-2fd451b1e0c7 · outbound

This paper cites Task-embedded control networks for few-shot imitation learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Task-embedded control networks for few-shot imitation learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.986558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.986558Z digest=sha256:3c1bfcc704406d63692a2004da6c79bf9a85666c38fef61c18170ba0bb8468fd

Observation 4e79cee5-c384-42e8-8fa6-e3c52dea7609 · outbound

This paper cites Language models are few-shot learners,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Language models are few-shot learners,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.989663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.989663Z digest=sha256:971627713e0fc9330d53daab90d295a9874fd01fb95dd0661d61e26a34836f3a

Observation b0233dc6-c36b-4f2d-b522-0825b6b43474 · outbound

This paper cites Flamingo: A visual language model for few- shot learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Flamingo: A visual language model for few- shot learning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.992047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.992047Z digest=sha256:9e8b61b809499be3dc154cb4eb064b57fd7d703a41659c9d5b53ba69df4fe546

Observation 724a961b-56b2-4e88-b090-13a0d7358b64 · outbound

This paper cites Prompting decision transformer for few-shot pol- icy generalization,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Prompting decision transformer for few-shot pol- icy generalization,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.995043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.995043Z digest=sha256:bfe732cd3351ab33878e45d8e972bbdad05fd0f67c3dae1a42b7793d4901d014

Observation c0b02a88-c2be-414f-b846-d0f8fbb38d4f · outbound

This paper cites AMAGO: Scalable In-Context Reinforcement Learning for Adaptive Agents.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos AMAGO: Scalable In-Context Reinforcement Learning for Adaptive Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.998019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.998019Z digest=sha256:9152e5c2bf1f9ee9c728ba9ffc0a88e539f7e855316d9171a448f632cf838b54

Observation 49b1f0ed-07a2-448f-97d6-684f2da3d7d8 · outbound

This paper cites In-context Reinforcement Learning with Algorithm Distillation.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos In-context Reinforcement Learning with Algorithm Distillation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.001004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.001004Z digest=sha256:058a1220830ac5e2274ef766f99b7341f983940fc9de5670516a77246f3d38e0

Observation d217f832-5e6c-455a-b125-f53f81695552 · outbound

This paper cites REGENT: A Retrieval-Augmented Generalist Agent That Can Act In-Context in New Environments.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos REGENT: A Retrieval-Augmented Generalist Agent That Can Act In-Context in New Environments

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.003802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.003802Z digest=sha256:efa0f414ac625c55615adc634309204b87f5064062b04c5c6c1f02b657339c25

Observation d1f3c36e-f9f0-42d6-9f23-d465ce58dc57 · outbound

This paper cites Keypoint Action Tokens Enable In-Context Imitation Learning in Robotics.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Keypoint Action Tokens Enable In-Context Imitation Learning in Robotics

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.007414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.007414Z digest=sha256:67e1d362f40806d2630c3b14685f63f9aef44499c124d1ee28ba7d819832d364

Observation e79e46de-9492-40d1-a858-2b00bc01fdc0 · outbound

This paper cites In-Context Imitation Learning via Next-Token Prediction.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos In-Context Imitation Learning via Next-Token Prediction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.010783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.010783Z digest=sha256:c7a2d7e7493b5c64e12e1b4b9752ed6b4e0c5c3c88e1d1aea52ef30fb1784963

Observation bc756b72-5d4d-4090-afa6-a22a82fce9ce · outbound

This paper cites Instant Policy: In-Context Imitation Learning via Graph Diffusion.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Instant Policy: In-Context Imitation Learning via Graph Diffusion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.013303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.013303Z digest=sha256:e0c94e89174eb383ba7c819c92bdea02a077f3568a64cb93ba0b819bb35b49fc

Observation 93d744c2-b19e-4a63-8727-6175ad7260bc · outbound

This paper cites Generalization to New Sequential Decision Making Tasks with In-Context Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Generalization to New Sequential Decision Making Tasks with In-Context Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.016697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.016697Z digest=sha256:2793ab8c2ed9fe8b7681fc6f31fef4ad80e8533424aa090b5c60646635b1c341

Observation 54e1324d-dd74-4240-b012-5a1fbc21dcaf · outbound

This paper cites Benchmarking General-Purpose In-Context Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Benchmarking General-Purpose In-Context Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.019940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.019940Z digest=sha256:19c6ea205e8f904ed068e64e7992b1ec5ca26807f91f9398caf74cff47d8c51f

Observation b0fb77fe-8fd0-4cb2-b171-cc76dc4ea49f · outbound

This paper cites General-Purpose In-Context Learning by Meta-Learning Transformers.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos General-Purpose In-Context Learning by Meta-Learning Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.022159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.022159Z digest=sha256:86cf0d52685f653724e6ce6f2205c87df637203645a7af0da7a22a1b153e9b57

Observation 6f9e3fb2-2f4b-4f51-8dde-23275b7f18c2 · outbound

This paper cites RICL: Adding In-Context Adaptability to Pre-Trained Vision-Language-Action Models.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos RICL: Adding In-Context Adaptability to Pre-Trained Vision-Language-Action Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.024764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.024764Z digest=sha256:681c82a2110a0b2acae598e108d10aae24dcc2daf99e7b97e4f420690ea14a01

Observation def31ae2-80fe-48f9-96b1-d8309a8ea97c · outbound

This paper cites Roboturk: A crowdsourcing platform for robotic skill learning through imitation,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Roboturk: A crowdsourcing platform for robotic skill learning through imitation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.027742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.027742Z digest=sha256:250f35f5699290369c868af32feab1d600533cb4b0ec39f29eb1f71f5800e70b

Observation 07f99e7f-6d36-4972-b129-f6ba770ffeb3 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.029863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.029863Z digest=sha256:92da115042510b013eb0236702c9a311b643197de93cfc7f82cc545e7e5e5503

Observation b74bd45c-f74f-4eaa-84d2-b3ffa5aaa2cd · outbound

This paper cites Droid: A large-scale in-the-wild robot manipulation dataset,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Droid: A large-scale in-the-wild robot manipulation dataset,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.031973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.031973Z digest=sha256:75c2b73508886ab650e2c98eeb9344443bafdc71575cc0a2d49b384b320d0237

Observation 8b1c8003-e080-485b-b741-df5d63ee743f · outbound

This paper cites MimicPlay: Long-Horizon Imitation Learning by Watching Human Play.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos MimicPlay: Long-Horizon Imitation Learning by Watching Human Play

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.033954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.033954Z digest=sha256:ab4126647dd54a9f1bd6fa05ade5d4476557f6c0c83677a9725c2e3cbce30e58

Observation 37bb429d-0435-4457-96ac-f0202c3c5f22 · outbound

This paper cites Learning latent plans from play,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Learning latent plans from play,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.037439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.037439Z digest=sha256:35e6b73a8378651f257958c8592c70c713f73d36e1fda1268d1443829397ce2b

Observation 9bf8e2bb-386b-424e-8cd8-5854ade764b1 · outbound

This paper cites MetaICL: Learning to Learn In Context.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos MetaICL: Learning to Learn In Context

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.039660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.039660Z digest=sha256:7316f2d6a5a1bf6593f0adf1087754a151f711c08ecc03411667160d26ce3da6

Observation dd95dc28-dced-43dd-8a9d-070730f29998 · outbound

This paper cites Okami: Teaching humanoid robots manipulation skills through single video imitation,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Okami: Teaching humanoid robots manipulation skills through single video imitation,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.042123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.042123Z digest=sha256:5f9432e4f2a565d11bc06840e095be04d67cae3d4ba0c540217694cd7462f9be

Observation 8a51b61e-29aa-4413-aa4d-e71152c903c6 · outbound

This paper cites Humanoid policy˜ human policy,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Humanoid policy˜ human policy,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.044696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.044696Z digest=sha256:ffa9ebb2a1c584719c4f917d930e20b71da64581d626db6f79ff54cb2a4558fc

Observation 509fba2d-f547-4d7d-a2b1-a95c9a41648b · outbound

This paper cites Representation and control of the task space in humans and humanoid robots,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Representation and control of the task space in humans and humanoid robots,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.046884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.046884Z digest=sha256:342f8f961930feb8da8ab3b09ef66a8943619a711369b282d52b676953b06719

Observation f88b31c8-61c2-44e0-8b32-f35e88c2554f · outbound

This paper cites Masked autoencoders are scalable vision learners,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Masked autoencoders are scalable vision learners,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.049590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.049590Z digest=sha256:fa4dfe6e50f718b97346e458696b16099711cad8d9f648af79607d3ac68b8c6e

Observation 3ec0d3bc-0cb9-44a5-ae09-2088c2739165 · outbound

This paper cites Rethinking Patch Dependence for Masked Autoencoders.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Rethinking Patch Dependence for Masked Autoencoders

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.051840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.051840Z digest=sha256:b3cd4cd7e1b540ef8b061828b52903134d346eb14c479c6b791ace5d5a91e2fe

Observation 10acbeeb-755b-47a5-817c-206621baabf8 · outbound

This paper cites Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.054289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.054289Z digest=sha256:f789818ca28d20a71a2b150eec3f0c9c310fc1fe40bb9540f6c664ded89e6176

Observation 40cd3cc1-a873-4d66-b080-b5d106956de5 · outbound

This paper cites Zero-Shot Robot Manipulation from Passive Human Videos.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Zero-Shot Robot Manipulation from Passive Human Videos

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.056679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.056679Z digest=sha256:4ef07530fc477e400eaa8782388ff3338a4069dd6552b6117fcee50b35fbdb63

Observation b60583e3-30e0-459a-849a-ee1c62674b06 · outbound

This paper cites Hand me the data: Fast robot adaptation via hand path retrieval,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Hand me the data: Fast robot adaptation via hand path retrieval,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.058929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.058929Z digest=sha256:39c7c206ae5d1265b25964bb0a7a8a23b843cfe031a665feaee6b4cefadfb30f

Observation 918742f8-31b4-4d7f-ad6f-a71874cdc860 · outbound

This paper cites Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.060965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.060965Z digest=sha256:5467e6d9f3ccdf4f63e9ba619a52c76453ddab989fc82e8e1d0821a4430b46b7

Observation 8ebeba5f-2ff0-44c0-b08a-a321f91e036d · outbound

This paper cites robosuite: A Modular Simulation Framework and Benchmark for Robot Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos robosuite: A Modular Simulation Framework and Benchmark for Robot Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.063023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.063023Z digest=sha256:8b8dd7ee68bad402c8702414d577f712c6ddf4eae958c48801e9474f6996ccc7

Observation 32148449-3222-4240-9882-2ea1883395bd · outbound

This paper cites Evolutionary principles in self-referential learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Evolutionary principles in self-referential learning,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.065798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.065798Z digest=sha256:c67db2e28adeff383ed1ca5d72ee6d2c68ec6a0c8b811dfd2a9d8ee51586d87f

Observation aee686bf-a4e2-4d51-80de-505bdfcf9e6b · outbound

This paper cites Meta-neural networks that learn by learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Meta-neural networks that learn by learning,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.068210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.068210Z digest=sha256:171340a20556fa6c97a2a4af026adac9340a76cefcac3592dd5430a347b55fd9

Observation fa7d62ca-4592-4742-9683-69c6a161def8 · outbound

This paper cites Meta-learning with memory-augmented neu- ral networks,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Meta-learning with memory-augmented neu- ral networks,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.070278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.070278Z digest=sha256:68e560778ed49b8b36b9bc906b67173bd9be19d8d2ded36e16dccb47c4dd75b3

Observation 02400fbd-9494-42c3-b527-6552e639eb97 · outbound

This paper cites Learning to learn using gradient descent,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Learning to learn using gradient descent,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.072563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.072563Z digest=sha256:a519b099c3f8bafe16d933afa3466819dd71915c5c2356d89148ca1a4e8a7d20

Observation adc6d6fb-deb2-42c8-9d04-e7115d8d3350 · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.074938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.074938Z digest=sha256:5da6c771bad58420e46e9e0d58c3225657df87e1c800200dcef441a366e6176d

Observation cdc1fc02-0533-4b80-9b16-a0a1c38eb02e · outbound

This paper cites Attention is all you need,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Attention is all you need,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.077854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.077854Z digest=sha256:6be0628658657c387552d77c7f6b63f01d892606bd3e83c8dcc7def41f50b3f7

Observation dca32268-79bc-46cc-8be0-8ab31de29413 · outbound

This paper cites RRL: Resnet as representation for Reinforcement Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos RRL: Resnet as representation for Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.079940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.079940Z digest=sha256:50dcc30717e8c191358c35aef054b8d8ae0f03eef68140f20f1c454fb98afedd

Observation c69ccdf7-8877-4ac6-ae8a-ee135e90f0c8 · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos R3M: A Universal Visual Representation for Robot Manipulation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.082371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.082371Z digest=sha256:aec56f4f3fc9f8175986538cb2bd909101b9038805b51c7cf2cead1c16e394c6

Observation 5d7d30ec-7118-44fe-994b-5689878dff2d · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelligence?.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Where are we in the search for an artificial visual cortex for embodied intelligence?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.085131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.085131Z digest=sha256:0bf608655ae4043efd613ee37f80c1ab07036087554098581d4b26ef1a7d750e

Observation 873910a3-fe4f-49e8-b716-9e1ff0868e50 · outbound

This paper cites Concept2robot: Learning manipulation concepts from instructions and human demonstrations,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Concept2robot: Learning manipulation concepts from instructions and human demonstrations,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.087661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.087661Z digest=sha256:e09e2184e66bbbb84066cad610ae834051f750dcf91728318e6d5310578c2139

Observation c4a10568-dce8-441f-8e13-4eb6d82b2fa4 · outbound

This paper cites Liv: Language-image representations and rewards for robotic control,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Liv: Language-image representations and rewards for robotic control,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.089935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.089935Z digest=sha256:f3472342708cb72e8903f3884ce5f52f5322d78b78e97e73d4cb70ca0ad164b2

Observation bb877225-36ee-4d3a-a5ee-bbf5e02b9119 · outbound

This paper cites VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.092177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.092177Z digest=sha256:da0f1f9ebdd38ad1340bb882a5b9e6e1f3e2051c6cc48e8fb35039a652ef3d39

Observation ab8183e0-9cea-4518-8edf-f37c2a33ba02 · outbound

This paper cites Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.094634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.094634Z digest=sha256:76b5488a8a917de0e1575e617e64388ec552f39574fbd0458db35695f820b6fb

Observation 2e30f510-7f36-47ca-a90a-90391e87da85 · outbound

This paper cites ScrewMimic: Bimanual Imitation from Human Videos with Screw Space Projection.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos ScrewMimic: Bimanual Imitation from Human Videos with Screw Space Projection

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.096917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.096917Z digest=sha256:71ec95e43e05b4832fcf790989a56f5562b37e333dc9fda77a04f9314673e88f

Observation 32091f57-4abb-4acf-95f9-b25915a68215 · outbound

This paper cites EgoMimic: Scaling Imitation Learning via Egocentric Video.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos EgoMimic: Scaling Imitation Learning via Egocentric Video

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.099335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.099335Z digest=sha256:ffd761cc23e2f2b0afaeb20ff58df4a873bb74fc7d10b1607ba5b0088c9de26c

Observation 7e090c30-8b6c-4e10-8bf9-d753907dd415 · outbound

This paper cites an unresolved cited work.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.101720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.101720Z digest=sha256:627c68b8dd7adff968ba05005796647741ac639f2e034c0f8a3f0f4dc08ff4cb

Observation 9fde1f81-89cf-4b44-bd30-56569c5c6e50 · outbound

This paper cites Reconstructing hands in 3D with transform- ers,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Reconstructing hands in 3D with transform- ers,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.103866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.103866Z digest=sha256:376168fc5f2a78da61d93246024e68687ae504a6c48ac69fc6d448fc3f7bf759

Observation 95d91406-94d2-4a22-a1e9-38d3cc1a2ba9 · outbound

This paper cites Phantom: Training Robots Without Robots Using Only Human Videos.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Phantom: Training Robots Without Robots Using Only Human Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.106113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.106113Z digest=sha256:e1c4dfeaf0577e79837ef337faf1e4f960b7068bd9e9a729821420b554b6a362

Observation f3066f7f-e8a0-4f0c-8bc1-bb9f7a5c211f · outbound

This paper cites Vision-based Manipulation from Single Human Video with Open-World Object Graphs.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.108784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.108784Z digest=sha256:88ee8c0615eff68905120c0ea2ed9e2c0b919d7adf973c1c7a16426a77b67ffd

Observation b7a1aacb-ddf0-4850-9ced-6162638e0d3f · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos DINOv2: Learning Robust Visual Features without Supervision

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.111093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.111093Z digest=sha256:28700848a5e051649129a29a9563c003e3beb4ac2e6288c3e11c52aee4cf5e4f

Observation c7959343-5e0f-4470-ab35-d99baa8b485a · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.113708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.113708Z digest=sha256:f40ceaa77ac0b932610e2993ed111c2a30e1013d842b4de586d8fe3e9ccceb78

Observation 629086a4-0c76-4702-99f1-dcef23c9f511 · outbound

This paper cites Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.116533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.116533Z digest=sha256:f0a3adba51bc44019982c0fa02c0d7f73220a49d3c9a2a8471a35a481e01f835

Observation c68f25b1-6ea0-4947-a80c-0f1efd5fdb79 · outbound

This paper cites Legato: Cross-embodiment imitation using a grasping tool,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Legato: Cross-embodiment imitation using a grasping tool,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.118520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.118520Z digest=sha256:656126a88462cac692b25f060ba8b856e6ade1c20f9569b06047f2b2ef65d37c

Observation b139d678-ebc3-4869-83b6-b94fe3e0e069 · outbound

This paper cites Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.121101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.121101Z digest=sha256:6eaf65e8ededb8984d0d6586795926d5545ca7831888085433172e06cca2be42

Observation 9c882efb-1139-4228-b6eb-e800c10490ce · outbound

This paper cites Does continual learning equally forget all pa- rameters?.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Does continual learning equally forget all pa- rameters?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.123316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.123316Z digest=sha256:9f37c6b20499126d8cd75803dcb65b61318d32d64ac04f0d5ec930311ec008b1

Observation eaca2272-ab9b-4419-9cba-fef227e19ff3 · outbound

This paper cites Segment anything,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Segment anything,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.125796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.125796Z digest=sha256:8c2a1edf5e3228a9de01a9bbb56fd4dbb6f879d7817401b31c438a40a8795a54

Observation 21fce94b-14a0-4ddf-a77e-23c89142ee12 · outbound

This paper cites Decoupling human and camera motion from videos in the wild,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Decoupling human and camera motion from videos in the wild,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.127656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.127656Z digest=sha256:84056f0060dc84034baa1458088294d73539e5a2a8b58182f8be574ea78d1f51

Pith citing papers

Observation 08184c78-6d50-4ee1-bff1-766e366bb783 · inbound

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations cites this paper.

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:40:33.752011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T00:37:12.170711Z digest=sha256:ce778fdb57579f1bed06433a83fcdd5238d7ef36c85d9408410766dc308e25b9

Observation 64b637e1-d756-4dcc-8b3d-8a34a03298fa · inbound

Bimanual Robot Manipulation via Multi-Agent In-Context Learning cites this paper.

Bimanual Robot Manipulation via Multi-Agent In-Context Learning MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:29:47.506260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T00:25:24.362191Z digest=sha256:f9332d8bfb80b1f88527a721fcadd1b37f65f44a675c7db6989cf4c4d2e91952

Observation f1cd0c19-d6c0-4d0f-92a5-81be84f8848f · inbound

HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos cites this paper.

HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:48.712142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T01:02:58.035784Z digest=sha256:5a159e736881b5f76e0ff6118196558de6ee4ba9c1309e9700cdb0cdfbd3dea7

Observation 46752ed2-8080-476f-9419-73bd2bdba9ab · inbound

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios cites this paper.

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:36:58.971265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T01:12:25.956015Z digest=sha256:109f24a21b2a038883e13f50d84a7ae587d6b40a87f19174835db92d7c6e84d9

Observation 9afd20f5-db4e-4edf-9feb-0023509c07bc · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.579213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T19:12:22.513577Z digest=sha256:e8b4e34c65a21213108aae11c97923ff2ee24c1902e5b0add8e3081732ef5b17

Observation 9da99683-5dd7-469a-8365-3812b0941822 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.621126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T05:11:07.089829Z digest=sha256:5f33ab6b9b82c35dec17c2720ebc499b40b491062f542f5796cbd30b00e3b75f

Observation 10c874f3-da37-429e-a71c-642627d80fd2 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T12:05:57.682386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:05:57.682386Z digest=sha256:5561232b5aed76a9be84e534039019655709db310e85599e134e3a017c7411f6

Observation 9f163656-9cde-4d67-a8db-f05d86f1c0ee · inbound

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time cites this paper.

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-09T22:16:36.338374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T22:16:31.529359Z digest=sha256:a6eab8ca10d7dfb811c8d2f06054153128ead4a61aeb8ee353042ba728a93787

Observation bca86fbd-c713-43c5-b318-7bf3c3740363 · inbound

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time cites this paper.

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T06:48:14.554799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:48:14.554799Z digest=sha256:18938ceda713700c372e6ead4cd5e149453e36c8313dffef045e53cd16e09100