UniMoFlow is a single flow-matching model for both text-to-motion generation and instruction-driven editing, trained with a 55,641-triplet synthetic dataset and a source-anchored sampling mode.
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages =
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
UniMoFlow: Grounding Instruction-Driven 3D Human Motion Editing in Generation
UniMoFlow is a single flow-matching model for both text-to-motion generation and instruction-driven editing, trained with a 55,641-triplet synthetic dataset and a source-anchored sampling mode.