Using a zero-phase DDSP vocoder as a differentiable resynthesizer after a low-cost neural predictor improves perceptual quality (DNSMOS) and intelligibility (STOI) on a DNS2020 subset, though causal large-model gains are mixed.
Genhancer: High-fidelity speech enhancement via generative modeling on discrete codec tokens,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Improving Resource-Efficient Speech Enhancement via Neural Differentiable DSP Vocoder Refinement
Using a zero-phase DDSP vocoder as a differentiable resynthesizer after a low-cost neural predictor improves perceptual quality (DNSMOS) and intelligibility (STOI) on a DNS2020 subset, though causal large-model gains are mixed.