Pith. sign in

REVIEW 2 cited by

Towards One Model to Rule All: Multilingual Strategy for Dialectal Code-Switching Arabic ASR

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2105.14779 v2 pith:O7MTD2XD submitted 2021-05-31 cs.CL cs.HCcs.SDeess.AS

classification cs.CLcs.HCcs.SDeess.AS
keywords arabicdialectalcode-switchingmonolingualmultilingualcharacterhandlinglanguage
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

With the advent of globalization, there is an increasing demand for multilingual automatic speech recognition (ASR), handling language and dialectal variation of spoken content. Recent studies show its efficacy over monolingual systems. In this study, we design a large multilingual end-to-end ASR using self-attention based conformer architecture. We trained the system using Arabic (Ar), English (En) and French (Fr) languages. We evaluate the system performance handling: (i) monolingual (Ar, En and Fr); (ii) multi-dialectal (Modern Standard Arabic, along with dialectal variation such as Egyptian and Moroccan); (iii) code-switching -- cross-lingual (Ar-En/Fr) and dialectal (MSA-Egyptian dialect) test cases, and compare with current state-of-the-art systems. Furthermore, we investigate the influence of different embedding/character representations including character vs word-piece; shared vs distinct input symbol per language. Our findings demonstrate the strength of such a model by outperforming state-of-the-art monolingual dialectal Arabic and code-switching Arabic ASR.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CAFE A Novel Code switching Dataset for Algerian Dialect French and English

    cs.SD 2024-11 conditional novelty 6.0 of 10

    CAFE is a new spontaneous speech corpus for Algerian dialect, French, and English code-switching, with 2.6 hours manually annotated and a Whisper benchmark reaching MER 0.310.

  2. Open Universal Arabic ASR Leaderboard

    cs.CL 2024-12 conditional novelty 5.0 of 10

    A new Arabic ASR leaderboard ranks 14 open-source models on five multi-dialect datasets and analyzes robustness, speaker bias, and efficiency.

Pith tools