Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Democratizing High-Fidelity Co-Speech Gesture Video Generation

cs.CV · 2025-07-09 · conditional · novelty 6.0

Co-speech gesture video is generated by first predicting 2D skeleton motion from audio via feature-concatenated diffusion, then rendering with an off-the-shelf video model, supported by a new 405-hour public dataset.

citing papers explorer

Showing 1 of 1 citing paper.

  • Democratizing High-Fidelity Co-Speech Gesture Video Generation cs.CV · 2025-07-09 · conditional · none · ref 7

    Co-speech gesture video is generated by first predicting 2D skeleton motion from audio via feature-concatenated diffusion, then rendering with an off-the-shelf video model, supported by a new 405-hour public dataset.