A Wav2Vec2 model pre-trained only on Dutch encodes Dutch phonetic and lexical information better than English-only or larger multilingual models, and this improvement carries over to Dutch ASR.
We use a sep- arate subset of MLS audiobook segments (held-out from w2v2- nl training data), as well as the IFADV corpus [23] of face- to-face conversational speech
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
What do self-supervised speech models know about Dutch? Analyzing advantages of language-specific pre-training
A Wav2Vec2 model pre-trained only on Dutch encodes Dutch phonetic and lexical information better than English-only or larger multilingual models, and this improvement carries over to Dutch ASR.