Audio-Mamba models outperform attention-based models such as WavLM and HuBERT on non-verbal emotion recognition, and a Renyi-divergence fusion approach called RENO further improves accuracy.
Emotion recognition in speech using mfcc and wavelet features,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Are Mamba-based Audio Foundation Models the Best Fit for Non-Verbal Emotion Recognition?
Audio-Mamba models outperform attention-based models such as WavLM and HuBERT on non-verbal emotion recognition, and a Renyi-divergence fusion approach called RENO further improves accuracy.