SB-RF integrates SB time marginals with RF velocity objectives to train a conditional velocity field that performs single Euler-step enhancement from noisy speech, showing competitive results on VoiceBank-DEMAND and superior results on low-SNR tests.
V oiceRestore: flow-matching transformers for speech recording quality restoration,
2 Pith papers cite this work. Polarity classification is still indexing.
representative citing papers
A new open-source pipeline creates a 5,078-hour Russian speech dataset with prosody annotations, and VITS and SEMamba models trained on it outperform those trained on existing Russian corpora under equalized budgets.
citing papers explorer
-
SB-RF: Schr\"odinger Bridge Rectified Flow for One-Step Robust Speech Enhancement
SB-RF integrates SB time marginals with RF velocity objectives to train a conditional velocity field that performs single Euler-step enhancement from noisy speech, showing competitive results on VoiceBank-DEMAND and superior results on low-SNR tests.
-
Balalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech
A new open-source pipeline creates a 5,078-hour Russian speech dataset with prosody annotations, and VITS and SEMamba models trained on it outperform those trained on existing Russian corpora under equalized budgets.