MoTe is a unified motion-text diffusion model that achieves strong text-to-motion generation and competitive motion captioning on HumanML3D and KIT by learning joint, conditional, and marginal distributions in one network.
Executing your commands via motion diffusion in latent space,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2024 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
MoTe: Learning Motion-Text Diffusion Model for Multiple Generation Tasks
MoTe is a unified motion-text diffusion model that achieves strong text-to-motion generation and competitive motion captioning on HumanML3D and KIT by learning joint, conditional, and marginal distributions in one network.