Pith. sign in

REVIEW 1 cited by

Modiff: Action-Conditioned 3D Motion Generation with Denoising Diffusion Probabilistic Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2301.03949 v2 pith:4IQ5R5SW submitted 2023-01-10 cs.CV

Modiff: Action-Conditioned 3D Motion Generation with Denoising Diffusion Probabilistic Models

classification cs.CV
keywords diffusionmotiongenerationmodelsprobabilisticaction-conditionedddpmdenoising
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Diffusion-based generative models have recently emerged as powerful solutions for high-quality synthesis in multiple domains. Leveraging the bidirectional Markov chains, diffusion probabilistic models generate samples by inferring the reversed Markov chain based on the learned distribution mapping at the forward diffusion process. In this work, we propose Modiff, a conditional paradigm that benefits from the denoising diffusion probabilistic model (DDPM) to tackle the problem of realistic and diverse action-conditioned 3D skeleton-based motion generation. We are a pioneering attempt that uses DDPM to synthesize a variable number of motion sequences conditioned on a categorical action. We evaluate our approach on the large-scale NTU RGB+D dataset and show improvements over state-of-the-art motion generation methods.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Kinetic Mining in Context: Few-Shot Action Synthesis via Text-to-Motion Distillation

    cs.CV 2025-12 conditional novelty 7.0

    A CLIP-guided teacher-student pipeline distills a text-to-motion prior into a few-shot action-to-motion generator, improving HAR top-1 accuracy by 23.1 points on 3 NTU-120 classes.