Pith. sign in

REVIEW 1 cited by

End-to-End Automatic Speech Recognition model for the Sudanese Dialect

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2212.10826 v1 pith:3UJDDGR5 submitted 2022-12-21 cs.CL

classification cs.CL
keywords dialectrecognitionmodelspeechsudaneseautomaticdatasetdesigning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Designing a natural voice interface rely mostly on Speech recognition for interaction between human and their modern digital life equipment. In addition, speech recognition narrows the gap between monolingual individuals to better exchange communication. However, the field lacks wide support for several universal languages and their dialects, while most of the daily conversations are carried out using them. This paper comes to inspect the viability of designing an Automatic Speech Recognition model for the Sudanese dialect, which is one of the Arabic Language dialects, and its complexity is a product of historical and social conditions unique to its speakers. This condition is reflected in both the form and content of the dialect, so this paper gives an overview of the Sudanese dialect and the tasks of collecting represented resources and pre-processing performed to construct a modest dataset to overcome the lack of annotated data. Also proposed end- to-end speech recognition model, the design of the model was formed using Convolution Neural Networks. The Sudanese dialect dataset would be a stepping stone to enable future Natural Language Processing research targeting the dialect. The designed model provided some insights into the current recognition task and reached an average Label Error Rate of 73.67%.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. BanglaDialecto: An End-to-End AI-Powered Regional Speech Standardization

    cs.CL 2024-11 conditional novelty 4.0 of 10

    An ASR + MT + TTS pipeline converts Noakhali dialect speech to standard Bangla, with Whisper-large V2 achieving 0.8% CER and BanglaT5 a 41.6 BLEU on the authors' NDD dataset.

Pith tools