A dual-contrastive disentanglement method factorizes videos into independent task and embodiment latents, then uses a parameter-efficient adapter on a frozen video diffusion model to synthesize robot executions from single human demonstrations without paired data.
Learning to transfer human hand skills for robot manipulations.arXiv preprint arXiv:2501.04169
2 Pith papers cite this work. Polarity classification is still indexing.
fields
cs.RO 2years
2026 2verdicts
UNVERDICTED 2representative citing papers
SynManDex generates human-like dexterous grasps for robots from synthetic human pre-grasps via retargeting and force-closure optimization, reporting 86.4% stability, 4.67/5 human-likeness, 80.7% sim success, and 83.3% real-robot success.
citing papers explorer
-
Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing
A dual-contrastive disentanglement method factorizes videos into independent task and embodiment latents, then uses a parameter-efficient adapter on a frozen video diffusion model to synthesize robot executions from single human demonstrations without paired data.
-
SynManDex: Synthesizing Human-like Dexterous Grasps from Synthetic Human Pre-Grasps
SynManDex generates human-like dexterous grasps for robots from synthetic human pre-grasps via retargeting and force-closure optimization, reporting 86.4% stability, 4.67/5 human-likeness, 80.7% sim success, and 83.3% real-robot success.