A fully automated pipeline generates audio-text pairs with grammar errors and disfluencies for spoken grammatical error correction, with four objective metrics to select the best generated data.
These metrics offer a systematic approach to selecting the optimal set of spoken augmented data from multiple candidates
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Data Augmentation for Spoken Grammatical Error Correction
A fully automated pipeline generates audio-text pairs with grammar errors and disfluencies for spoken grammatical error correction, with four objective metrics to select the best generated data.