On a roughly six-hour, three-dialect Common Voice subset, MFCC features with a CNN outperform wavelet features with an RNN by about 25 accuracy points.
Convolutional neural networks for speech recognition
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Hybrid Deep Learning and Signal Processing for Arabic Dialect Recognition in Low-Resource Settings
On a roughly six-hour, three-dialect Common Voice subset, MFCC features with a CNN outperform wavelet features with an RNN by about 25 accuracy points.