paper-with-me

Papers

Evaluation of Adversarial Robustness in Arabic Language Models

2026-07-28 · Anwar Alajmi, Ayed Salman, Imtiaz Ahmad arxiv

The emergence of the recent outstanding capabilities of Arabic Language Models has opened doors for exposing their vulnerabilities. One of the major security risks associated with such Natural Language Processing models is adversarial attacks. These attacks can deceive the model into the wrong prediction, raising critical model security and safety concerns. This study aims to assess the robustness of five state-of-the-art Arabic Language Models under a distinct set of Arabic adversarial attacks applied at various levels of granularity and using different example generation strategies. We also explore a defense technique based on adversarial training to enhance model robustness. The results show that insertion of diacritics can reduce the accuracy of some models by 92% while maintaining a low perturbation distance. For word-level attacks, manipulating Arabic conjunctions preserves high semantic similarity scores, low perturbation distance, and leads to an accuracy degradation of up to 58%. For sentence-level attacks, paraphrasing proves its effectiveness by an average reduction of 76% in the victim models' performance. While adversarial training improves overall resilience, with MARBERT being the most robust and AraBERT showing the greatest relative gains, challenges persist, particularly against character-level noise. These findings highlight both the potential and limitations of current defense strategies in morphologically rich languages like Arabic.

📄 PDF Abstract BibTeX arXiv:2607.25814

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial RobustnessSemantic Similarity

Similar Papers 제목 키워드 기반

Arabic Synonym BERT-based Adversarial Examples for Text Classification

2024-02-05 · Norah Alshahrani, Saied Alshahrani, Esma Wali, Jeanna Matthews

Text classification systems have been proven vulnerable to adversarial text examples, modified versions of the original text examples that are often unnoticed by human eyes, yet can force text classification models to al…

Adversarial TextLanguage ModelingLanguage ModellingMasked Language Modeling+2

Towards stable AI systems for Evaluating Arabic Pronunciations

2025-08-27 · Hadi Zaatiti, Hatem Hajri, Osama Abdullah, Nader Masmoudi arxiv

Modern Arabic ASR systems such as wav2vec 2.0 excel at word- and sentence-level transcription, yet struggle to classify isolated letters. In this study, we show that this phoneme-level task, crucial for language learning…

Improved Generalization of Arabic Text Classifiers

2019-08-01 · WS 2019 8 · Alaa Khaddaj, Hazem Hajj, Wassim El-Hajj

While transfer learning for text has been very active in the English language, progress in Arabic has been slow, including the use of Domain Adaptation (DA). Domain Adaptation is used to generalize the performance of any…

Domain AdaptationTransfer Learning

Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification

2025-05-30 · Badr M. Abdullah, Matthew Baas, Bernd Möbius, Dietrich Klakow

Arabic dialect identification (ADI) systems are essential for large-scale data collection pipelines that enable the development of inclusive speech technologies for Arabic language varieties. However, the reliability of …

Dialect IdentificationVoice Conversion

N-Shot Benchmarking of Whisper on Diverse Arabic Speech Recognition

2023-06-05 · Bashar Talafha, Abdul Waheed, Muhammad Abdul-Mageed

Whisper, the recently developed multilingual weakly supervised model, is reported to perform well on multiple speech recognition benchmarks in both monolingual and multilingual settings. However, it is not clear how Whis…

Arabic Speech RecognitionBenchmarkingspeech-recognitionSpeech Recognition