paper-with-me

홈 › Papers

VoxArabica: A Robust Dialect-Aware Arabic Speech Recognition System

2023-10-17 · Abdul Waheed, Bashar Talafha, Peter Sullivan, AbdelRahim Elmadany, Muhammad Abdul-Mageed

Arabic is a complex language with many varieties and dialects spoken by over 450 millions all around the world. Due to the linguistic diversity and variations, it is challenging to build a robust and generalized ASR system for Arabic. In this work, we address this gap by developing and demoing a system, dubbed VoxArabica, for dialect identification (DID) as well as automatic speech recognition (ASR) of Arabic. We train a wide range of models such as HuBERT (DID), Whisper, and XLS-R (ASR) in a supervised setting for Arabic DID and ASR tasks. Our DID models are trained to identify 17 different dialects in addition to MSA. We finetune our ASR models on MSA, Egyptian, Moroccan, and mixed data. Additionally, for the remaining dialects in ASR, we provide the option to choose various models such as Whisper and MMS in a zero-shot setting. We integrate these models into a single web interface with diverse features such as audio recording, file upload, model selection, and the option to raise flags for incorrect outputs. Overall, we believe VoxArabica will be useful for a wide range of audiences concerned with Arabic research. Our system is currently running at https://cdce-206-12-100-168.ngrok.io/.

📄 PDF Abstract BibTeX arXiv:2310.11069

Code (0)

등록된 구현이 없습니다.

Tasks

Arabic Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect IdentificationDiversityModel Selectionspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect

2025-04-03 · Hedi Naouara, Jean-Pierre Lorré, Jérôme Louradour

Developing Automatic Speech Recognition (ASR) systems for Tunisian Arabic Dialect is challenging due to the dialect's linguistic complexity and the scarcity of annotated speech datasets. To address these challenges, we p…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+2

Development of a TV Broadcasts Speech Recognition System for Qatari Arabic

2014-05-01 · LREC 2014 5 · Mohamed Elmahdy, Mark Hasegawa-Johnson, Eiman Mustafawi

A major problem with dialectal Arabic speech recognition is due to the sparsity of speech resources. In this paper, a transfer learning framework is proposed to jointly use a large amount of Modern Standard Arabic (MSA) …

Arabic Speech RecognitionLanguage ModelingLanguage Modellingspeech-recognition+2

NADI 2025: The First Multidialectal Arabic Speech Processing Shared Task

2025-09-02 · Bashar Talafha, Hawau Olamide Toyin, Peter Sullivan, AbdelRahim Elmadany 외 arxiv

We present the findings of the sixth Nuanced Arabic Dialect Identification (NADI 2025) Shared Task, which focused on Arabic speech dialect processing across three subtasks: spoken dialect identification (Subtask 1), spee…

Speech Recognition

Speech Recognition Challenge in the Wild: Arabic MGB-3

2017-09-21 · Ahmed Ali, Stephan Vogel, Steve Renals

This paper describes the Arabic MGB-3 Challenge - Arabic Speech Recognition in the Wild. Unlike last year's Arabic MGB-2 Challenge, for which the recognition task was based on more than 1,200 hours broadcast TV news reco…

Arabic Speech RecognitionDialect Identificationspeech-recognitionSpeech Recognition

A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain

2024-03-07 · Qusai Abo Obaidah, Muhy Eddin Za'ter, Adnan Jaljuli, Ali Mahboub 외

This work is an attempt to introduce a comprehensive benchmark for Arabic speech recognition, specifically tailored to address the challenges of telephone conversations in Arabic language. Arabic, characterized by its ri…

Arabic Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversity+2