Towards One Model to Rule All: Multilingual Strategy for Dialectal Code-Switching Arabic ASR

Ahmed Abdelali; Ahmed Ali; Amir Hussein; Shammur Absar Chowdhury

arxiv: 2105.14779 · v2 · pith:O7MTD2XDnew · submitted 2021-05-31 · 💻 cs.CL · cs.HC· cs.SD· eess.AS

Towards One Model to Rule All: Multilingual Strategy for Dialectal Code-Switching Arabic ASR

Shammur Absar Chowdhury , Amir Hussein , Ahmed Abdelali , Ahmed Ali This is my paper

classification 💻 cs.CL cs.HCcs.SDeess.AS

keywords arabicdialectalcode-switchingmonolingualmultilingualcharacterhandlinglanguage

0 comments

read the original abstract

With the advent of globalization, there is an increasing demand for multilingual automatic speech recognition (ASR), handling language and dialectal variation of spoken content. Recent studies show its efficacy over monolingual systems. In this study, we design a large multilingual end-to-end ASR using self-attention based conformer architecture. We trained the system using Arabic (Ar), English (En) and French (Fr) languages. We evaluate the system performance handling: (i) monolingual (Ar, En and Fr); (ii) multi-dialectal (Modern Standard Arabic, along with dialectal variation such as Egyptian and Moroccan); (iii) code-switching -- cross-lingual (Ar-En/Fr) and dialectal (MSA-Egyptian dialect) test cases, and compare with current state-of-the-art systems. Furthermore, we investigate the influence of different embedding/character representations including character vs word-piece; shared vs distinct input symbol per language. Our findings demonstrate the strength of such a model by outperforming state-of-the-art monolingual dialectal Arabic and code-switching Arabic ASR.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

When Multiple Scripts Matter: Evaluating ASR in Clinical Settings
cs.CL 2026-06 unverdicted novelty 6.0

MultiClin benchmark shows multiscript-aware evaluation is fairer than single-reference metrics for clinical ASR, and script unification during training yields the best performance.