A HuBERT and HiFi-GAN pipeline with f0 features is proposed for unified voice and accent conversion in speech and singing, but the claimed gains are not backed by reproducible evidence.
V oice Conversion Using Speech-to-Speech Neuro-Style Transfer,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.SD 1years
2024 1verdicts
REJECT 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
A Unified Model For Voice and Accent Conversion In Speech and Singing using Self-Supervised Learning and Feature Extraction
A HuBERT and HiFi-GAN pipeline with f0 features is proposed for unified voice and accent conversion in speech and singing, but the claimed gains are not backed by reproducible evidence.