A modern Conformer ASR model's predictions are tied to vowel formants (F1 and F2), sibilant fricative spectral peaks, and plosive release bursts, per SPES feature attributions on TIMIT.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution
A modern Conformer ASR model's predictions are tied to vowel formants (F1 and F2), sibilant fricative spectral peaks, and plosive release bursts, per SPES feature attributions on TIMIT.