By turning each layer's width into a continuous, noise-smoothed parameter, the authors train speech models whose sizes shrink during training, reducing FLOPs and size by roughly 80–90% in their case studies.
Title resolution pending
1 Pith paper cite this work, alongside 12 external citations. Polarity classification is still indexing.
1
Pith paper citing it
12
external citations · OpenAlex
fields
cs.SD 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Performance and Complexity Trade-off Optimization of Speech Models During Training
By turning each layer's width into a continuous, noise-smoothed parameter, the authors train speech models whose sizes shrink during training, reducing FLOPs and size by roughly 80–90% in their case studies.