Pith. sign in

REVIEW 1 cited by

Generation of Complex 3D Human Motion by Temporal and Spatial Composition of Diffusion Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.11920 v1 pith:SAYSAPM6 submitted 2024-09-18 cs.CV cs.LG

classification cs.CVcs.LG
keywords complexhumanmotiondiffusionduringmodelsmovementstraining
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

In this paper, we address the challenge of generating realistic 3D human motions for action classes that were never seen during the training phase. Our approach involves decomposing complex actions into simpler movements, specifically those observed during training, by leveraging the knowledge of human motion contained in GPTs models. These simpler movements are then combined into a single, realistic animation using the properties of diffusion models. Our claim is that this decomposition and subsequent recombination of simple movements can synthesize an animation that accurately represents the complex input action. This method operates during the inference phase and can be integrated with any pre-trained diffusion model, enabling the synthesis of motion classes not present in the training data. We evaluate our method by dividing two benchmark human motion datasets into basic and complex actions, and then compare its performance against the state-of-the-art.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CoMA: Compositional Human Motion Generation with Multi-modal Agents

    cs.CV 2024-12 conditional novelty 6.0 of 10

    CoMA generates and edits 3D human motion from complex text by chaining GPT-4o planning, body-part-specific masked transformer generation, and VLM-based self-correction.

Pith tools