Stepback trains a voice converter with two decoders and a self-destructive loss to separate speaker identity from linguistic content, but the preprint contains no reported evaluation results.
Preparatory Stage For simplicity, we denote the content encoder input (source speech), the speaker identity, and the converted speech with x, i, y ϵ X, I, Y , respectively
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SD 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Stepback: Enhanced Disentanglement for Voice Conversion via Multi-Task Learning
Stepback trains a voice converter with two decoders and a self-destructive loss to separate speaker identity from linguistic content, but the preprint contains no reported evaluation results.