Pith. sign in

REVIEW 5 cited by

Learning to Generate Diverse Dance Motions with Transformer

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2008.08171 v1 pith:KOS5LCIR submitted 2020-08-18 cs.CV cs.GR

classification cs.CVcs.GR
keywords dancemotioncomplexgenerateintroducemotionssystemcapture
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

With the ongoing pandemic, virtual concerts and live events using digitized performances of musicians are getting traction on massive multiplayer online worlds. However, well choreographed dance movements are extremely complex to animate and would involve an expensive and tedious production process. In addition to the use of complex motion capture systems, it typically requires a collaborative effort between animators, dancers, and choreographers. We introduce a complete system for dance motion synthesis, which can generate complex and highly diverse dance sequences given an input music sequence. As motion capture data is limited for the range of dance motions and styles, we introduce a massive dance motion data set that is created from YouTube videos. We also present a novel two-stream motion transformer generative model, which can generate motion sequences with high flexibility. We also introduce new evaluation metrics for the quality of synthesized dance motions, and demonstrate that our system can outperform state-of-the-art methods. Our system provides high-quality animations suitable for large crowds for virtual concerts and can also be used as reference for professional animation pipelines. Most importantly, we show that vast online videos can be effective in training dance motion models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Text Dictates, Music Decorates: Energy-based Attention for Editable Dance Motion Generation

    cs.AI 2026-06 unverdicted novelty 7.0 of 10

    STREAM decouples text (via AdaLN) from music (via energy-based BEAM attention) to generate editable, musically aligned dance motions with a new annotated dataset and editability metric.

  2. Spatial-Temporal Graph Mamba for Music-Guided Dance Video Synthesis

    cs.CV 2025-07 conditional novelty 6.0 of 10

    STG-Mamba generates dance videos from music using a spatial-temporal graph Mamba block for skeleton generation and forward-backward self-supervised losses for video synthesis, reporting SOTA on benchmarks.

  3. Stochastic Human Motion Prediction with Memory of Action Transition and Action Characteristic

    cs.CV 2025-07 conditional novelty 6.0 of 10

    Adding a soft-transition action bank, an action characteristic bank, and adaptive attention fusion to the WAT baseline improves action-conditioned human motion prediction on four benchmarks.

  4. DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling

    cs.GR 2025-06 conditional novelty 6.0 of 10

    DuetGen is a two-stage masked-modeling system that converts music into synchronized, interactive two-person dance motion, and claims state-of-the-art results on the DD100 duet dataset.

  5. MotionRAG-Diff: A Retrieval-Augmented Diffusion Framework for Long-Term Music-to-Dance Generation

    cs.SD 2025-06 reject novelty 4.0 of 10

    MotionRAG-Diff combines contrastive retrieval from a motion graph with a diffusion model to generate long music-synchronized dance, reporting strong beat alignment but mixed quality scores.

Pith tools