Multiple choice learning matches permutation invariant training for speech separation on WSJ0-mix and LibriMix with up to 20 speakers, at lower loss-computation cost.
Annealed Multiple Choice Learning: Overcoming limitations of Winner-takes-all with annealing
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
We introduce Annealed Multiple Choice Learning (aMCL) which combines simulated annealing with MCL. MCL is a learning framework handling ambiguous tasks by predicting a small set of plausible hypotheses. These hypotheses are trained using the Winner-takes-all (WTA) scheme, which promotes the diversity of the predictions. However, this scheme may converge toward an arbitrarily suboptimal local minimum, due to the greedy nature of WTA. We overcome this limitation using annealing, which enhances the exploration of the hypothesis space during training. We leverage insights from statistical physics and information theory to provide a detailed description of the model training trajectory. Additionally, we validate our algorithm by extensive experiments on synthetic datasets, on the standard UCI benchmark, and on speech separation.
fields
cs.SD 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Multiple Choice Learning for Efficient Speech Separation with Many Speakers
Multiple choice learning matches permutation invariant training for speech separation on WSJ0-mix and LibriMix with up to 20 speakers, at lower loss-computation cost.