CTFT-Net jointly reconstructs magnitude and phase for speech bandwidth extension and reports lower log-spectral distance than NU-Wave, WSRGlow, NVSR, and AERO, but its own tables and core equation contain inconsistencies.
It shows strong performance across a wide range of input sampling rates ranging from 2 kHz to 48 kHz
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SD 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
A High-Fidelity Speech Super Resolution Network using a Complex Global Attention Module with Spectro-Temporal Loss
CTFT-Net jointly reconstructs magnitude and phase for speech bandwidth extension and reports lower log-spectral distance than NU-Wave, WSRGlow, NVSR, and AERO, but its own tables and core equation contain inconsistencies.