An inference-time module uses ASR alignments and per-segment noisy/enhanced mixing to reduce over-suppression in speech enhancement outputs.
Models and datasets We evaluate our approach using three state-of-the-art SE models, including CMGAN [1], SEMamba [2], and SGMSE [3]
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
baseline 1
citation-polarity summary
fields
eess.AS 1years
2026 1verdicts
CONDITIONAL 1roles
baseline 1polarities
baseline 1representative citing papers
citing papers explorer
-
Mitigating Over-Suppression in Speech Enhancement via Inference-Time Rethink-and-Refine Correction Module
An inference-time module uses ASR alignments and per-segment noisy/enhanced mixing to reduce over-suppression in speech enhancement outputs.