Using Whisper transcriptions plus DeBERTa text features, a lexical-only pipeline beats acoustic-only models on MELD speech emotion recognition: 51.5% vs 49.3% weighted F1.
Survey on bimodal speech emotion recognition from acoustic and linguistic informa- tion fusion,
1 Pith paper cite this work, alongside 106 external citations. Polarity classification is still indexing.
1
Pith paper citing it
106
external citations · OpenAlex
citation-role summary
background 1
citation-polarity summary
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
On the Contribution of Lexical Features to Speech Emotion Recognition
Using Whisper transcriptions plus DeBERTa text features, a lexical-only pipeline beats acoustic-only models on MELD speech emotion recognition: 51.5% vs 49.3% weighted F1.