paper-with-me

홈 › Papers

Towards stable AI systems for Evaluating Arabic Pronunciations

2025-08-27 · Hadi Zaatiti, Hatem Hajri, Osama Abdullah, Nader Masmoudi arxiv

Modern Arabic ASR systems such as wav2vec 2.0 excel at word- and sentence-level transcription, yet struggle to classify isolated letters. In this study, we show that this phoneme-level task, crucial for language learning, speech therapy, and phonetic research, is challenging because isolated letters lack co-articulatory cues, provide no lexical context, and last only a few hundred milliseconds. Recogniser systems must therefore rely solely on variable acoustic cues, a difficulty heightened by Arabic's emphatic (pharyngealized) consonants and other sounds with no close analogues in many languages. This study introduces a diverse, diacritised corpus of isolated Arabic letters and demonstrates that state-of-the-art wav2vec 2.0 models achieve only 35% accuracy on it. Training a lightweight neural network on wav2vec embeddings raises performance to 65%. However, adding a small amplitude perturbation (epsilon = 0.05) cuts accuracy to 32%. To restore robustness, we apply adversarial training, limiting the noisy-speech drop to 9% while preserving clean-speech accuracy. We detail the corpus, training pipeline, and evaluation protocol, and release, on demand, data and code for reproducibility. Finally, we outline future work extending these methods to word- and sentence-level frameworks, where precise letter pronunciation remains critical.

📄 PDF Abstract BibTeX arXiv:2508.19587

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

An ensemble-based framework for mispronunciation detection of Arabic phonemes

2023-01-03 · Sukru Selim Calik, Ayhan Kucukmanisa, Zeynep Hilal Kilimci

Determination of mispronunciations and ensuring feedback to users are maintained by computer-assisted language learning (CALL) systems. In this work, we introduce an ensemble model that defines the mispronunciation of Ar…

Ensemble Learning

Anomaly detection with a variational autoencoder for Arabic mispronunciation detection

2024-06-25 · International Journal of Speech Technology 2024 6 · Meriem Lounis, Bilal Dendani, Halima Bahi

Computer-assisted language learning (CALL) systems increasingly arouse a significant interest and establish a presence in automated foreign language learning. They enhance traditional learning methods by providing acces…

Anomaly Detection

LAMAD: A Linguistic Attentional Model for Arabic Text Diacritization

2021-11-01 · Findings (EMNLP) 2021 11 · Raeed Al-Sabri, Jianliang Gao

In Arabic Language, diacritics are used to specify meanings as well as pronunciations. However, diacritics are often omitted from written texts, which increases the number of possible meanings and pronunciations. This le…

Arabic Text Diacritization

A Multitask Learning Approach for Diacritic Restoration

2020-06-07 · ACL 2020 6 · Sawsan Alqahtani, Ajay Mishra, Mona Diab

In many languages like Arabic, diacritics are used to specify pronunciations as well as meanings. Such diacritics are often omitted in written text, increasing the number of possible pronunciations and meanings for a wor…

Multi-Task LearningPart-Of-Speech Tagging

Developing LMF-XML Bilingual Dictionaries for Colloquial Arabic Dialects

2012-05-01 · LREC 2012 5 · David Graff, Mohamed Maamouri

The Linguistic Data Consortium and Georgetown University Press are collaborating to create updated editions of bilingual diction- aries that had originally been published in the 1960's for English-speaking learners of Mo…

Management