paper-with-me

홈 › Papers

Open Universal Arabic ASR Leaderboard

2024-12-18 · Yingzhi Wang, Anas Alhmoud, Muhammad Alqurishi

In recent years, the enhanced capabilities of ASR models and the emergence of multi-dialect datasets have increasingly pushed Arabic ASR model development toward an all-dialect-in-one direction. This trend highlights the need for benchmarking studies that evaluate model performance on multiple dialects, providing the community with insights into models' generalization capabilities. In this paper, we introduce Open Universal Arabic ASR Leaderboard, a continuous benchmark project for open-source general Arabic ASR models across various multi-dialect datasets. We also provide a comprehensive analysis of the model's robustness, speaker adaptation, inference efficiency, and memory consumption. This work aims to offer the Arabic ASR community a reference for models' general performance and also establish a common evaluation framework for multi-dialectal Arabic ASR models.

📄 PDF Abstract BibTeX arXiv:2412.13788

Code (1)

Natural-Language-Processing-Elm/open_universal_arabic_asr_leaderboard 공식 구현 pytorch

Tasks

Benchmarking

Similar Papers 제목 키워드 기반

State-of-the-Art Arabic Language Modeling with Sparse MoE Fine-Tuning and Chain-of-Thought Distillation

2026-04-07 · Navan Preet Singh, Anurag Garikipati, Ahmed Abulkhair, Jyani Akshay Jagdishbhai 외 arxiv

This paper introduces Arabic-DeepSeek-R1, an application-driven open-source Arabic LLM that leverages a sparse MoE backbone to address the digital equity gap for under-represented languages, and establishes a new SOTA ac…

CamelEval: Advancing Culturally Aligned Arabic Language Models and Benchmarks

2024-09-19 · Zhaozhi Qian, Faroq Altam, Muhammad Alqurishi, Riad Souissi

Large Language Models (LLMs) are the cornerstones of modern artificial intelligence systems. This paper introduces Juhaina, a Arabic-English bilingual LLM specifically designed to align with the values and preferences of…

Instruction FollowingOpen-Ended Question AnsweringQuestion Answering

Open Automatic Speech Recognition Models for Classical and Modern Standard Arabic

2025-07-18 · Lilit Grigoryan, Nikolay Karpov, Enas Albasiri, Vitaly Lavrukhin 외 arxiv

Despite Arabic being one of the most widely spoken languages, the development of Arabic Automatic Speech Recognition (ASR) systems faces significant challenges due to the language's complexity, and only a limited number …

Speech Recognition

3rd Place Solution for Google Universal Image Embedding

2022-10-14 · Nobuaki Aoki, Yasumasa Namba

This paper presents the 3rd place solution to the Google Universal Image Embedding Competition on Kaggle. We use ViT-H/14 from OpenCLIP for the backbone of ArcFace, and trained in 2 stage. 1st stage is done with freezed …

Universal Dependencies for Arabic

2017-04-01 · WS 2017 4 · Dima Taji, Nizar Habash, Daniel Zeman

We describe the process of creating NUDAR, a Universal Dependency treebank for Arabic. We present the conversion from the Penn Arabic Treebank to the Universal Dependency syntactic representation through an intermediate …

Machine TranslationQuestion Answering