A speech-text cross-attention model improves dysarthria detection and severity classification over speech-only models on the UA-Speech database, but the gain disappears for detection on unseen speakers and words.
Spectro-temporal representation of speech for intelligibility assessment of dysarthria,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
method 1
citation-polarity summary
fields
cs.AI 1years
2024 1verdicts
CONDITIONAL 1roles
method 1polarities
use method 1representative citing papers
citing papers explorer
-
A Multi-modal Approach to Dysarthria Detection and Severity Assessment Using Speech and Text Information
A speech-text cross-attention model improves dysarthria detection and severity classification over speech-only models on the UA-Speech database, but the gain disappears for detection on unseen speakers and words.