paper-with-me

홈 › Papers

Arabic Speech Recognition by End-to-End, Modular Systems and Human

2021-01-21 · Amir Hussein, Shinji Watanabe, Ahmed Ali

Recent advances in automatic speech recognition (ASR) have achieved accuracy levels comparable to human transcribers, which led researchers to debate if the machine has reached human performance. Previous work focused on the English language and modular hidden Markov model-deep neural network (HMM-DNN) systems. In this paper, we perform a comprehensive benchmarking for end-to-end transformer ASR, modular HMM-DNN ASR, and human speech recognition (HSR) on the Arabic language and its dialects. For the HSR, we evaluate linguist performance and lay-native speaker performance on a new dataset collected as a part of this study. For ASR the end-to-end work led to 12.5%, 27.5%, 33.8% WER; a new performance milestone for the MGB2, MGB3, and MGB5 challenges respectively. Our results suggest that human performance in the Arabic language is still considerably better than the machine with an absolute WER gap of 3.5% on average.

📄 PDF Abstract BibTeX arXiv:2101.08454

Code (1)

espnet/espnet/tree/master/egs/mgb2/asr1 공식 구현 pytorch

Tasks

Arabic Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Benchmarkingspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Advancing Arabic Speech Recognition Through Large-Scale Weakly Supervised Learning

2025-04-16 · Mahmoud Salhab, Marwan Elghitany, Shameed Sait, Syed Sibghat Ullah 외

Automatic speech recognition (ASR) is crucial for human-machine interaction in diverse applications like conversational agents, industrial robotics, call center automation, and automated subtitling. However, developing h…

Arabic Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+2

A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain

2024-03-07 · Qusai Abo Obaidah, Muhy Eddin Za'ter, Adnan Jaljuli, Ali Mahboub 외

This work is an attempt to introduce a comprehensive benchmark for Arabic speech recognition, specifically tailored to address the challenges of telephone conversations in Arabic language. Arabic, characterized by its ri…

Arabic Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversity+2

LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect

2025-04-03 · Hedi Naouara, Jean-Pierre Lorré, Jérôme Louradour

Developing Automatic Speech Recognition (ASR) systems for Tunisian Arabic Dialect is challenging due to the dialect's linguistic complexity and the scarcity of annotated speech datasets. To address these challenges, we p…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+2

German-Arabic Speech-to-Speech Translation for Psychiatric Diagnosis

2020-12-01 · COLING (WANLP) 2020 12 · Juan Hussain, Mohammed Mediani, Moritz Behr, M. Amin Cheragui 외

In this paper we present the natural language processing components of our German-Arabic speech-to-speech translation system which is being deployed in the context of interpretation during psychiatric, diagnostic intervi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderDiagnostic+7

CV-18 NER: Augmented Common Voice for Named Entity Recognition from Arabic Speech

2026-04-02 · Youssef Saidi, Haroun Elleuch, Fethi Bougares arxiv

End-to-end speech Named Entity Recognition (NER) aims to directly extract entities from speech. Prior work has shown that end-to-end (E2E) approaches can outperform cascaded pipelines for English, French, and Chinese, bu…