Pith. sign in

Text-to-audio generation using instruction guided latent diffusion model,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.SD 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

support 1

representative citing papers

Recomposer: Event-roll-guided generative audio editing

cs.SD · 2025-09-05 · conditional · novelty 6.0

An encoder-decoder transformer conditions on text actions and a time-aligned event roll to delete, insert, or enhance individual sound events in dense audio scenes.

citing papers explorer

Showing 1 of 1 citing paper.

  • Recomposer: Event-roll-guided generative audio editing cs.SD · 2025-09-05 · conditional · none · ref 5

    An encoder-decoder transformer conditions on text actions and a time-aligned event roll to delete, insert, or enhance individual sound events in dense audio scenes.