A two-stage Whisper-plus-FlanT5 system with diversity-based hypothesis selection reduces dysarthric speech WER from 11.60% to 7.34% on the development set, while single-word recognition stays at 63.08% WER.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Exploring Generative Error Correction for Dysarthric Speech Recognition
A two-stage Whisper-plus-FlanT5 system with diversity-based hypothesis selection reduces dysarthric speech WER from 11.60% to 7.34% on the development set, while single-word recognition stays at 63.08% WER.