A two-stage inpainting-based video diffusion transformer that reuses pretrained attention to reenact hand-object interactions with novel objects, reporting SOTA performance on Re-HOLD and a new in-the-wild dataset.
Cg-hoi: Contact-guided 3d human-object interaction generation
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.GR 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
iDiT-HOI: Inpainting-based Hand Object Interaction Reenactment via Video Diffusion Transformer
A two-stage inpainting-based video diffusion transformer that reuses pretrained attention to reenact hand-object interactions with novel objects, reporting SOTA performance on Re-HOLD and a new in-the-wild dataset.