DnR-nonverbal moves non-verbal vocal sounds into the speech stem for cinematic audio source separation, fixing the misallocation of laughter and screams to the effect stem.
Motivation In the actual movie audio, we can decomposex s as follows: xs =x v +x n,(2) wherex v andx n correspond to the waveforms of verbal and non-verbal sounds, respectively
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
dataset 1
citation-polarity summary
fields
cs.SD 1years
2025 1verdicts
CONDITIONAL 1roles
dataset 1polarities
background 1representative citing papers
citing papers explorer
-
DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds
DnR-nonverbal moves non-verbal vocal sounds into the speech stem for cinematic audio source separation, fixing the misallocation of laughter and screams to the effect stem.