On a roughly six-hour, three-dialect Common Voice subset, MFCC features with a CNN outperform wavelet features with an RNN by about 25 accuracy points.
https://openslr.org
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Hybrid Deep Learning and Signal Processing for Arabic Dialect Recognition in Low-Resource Settings
On a roughly six-hour, three-dialect Common Voice subset, MFCC features with a CNN outperform wavelet features with an RNN by about 25 accuracy points.