paper-with-me

홈 › Papers

Voice of a Continent: Mapping Africa's Speech Technology Frontier

2025-05-24 · AbdelRahim Elmadany, Sang Yun Kwon, Hawau Olamide Toyin, Alcides Alcoba Inciarte, Hanan Aldarmaki, Muhammad Abdul-Mageed

Africa's rich linguistic diversity remains significantly underrepresented in speech technologies, creating barriers to digital inclusion. To alleviate this challenge, we systematically map the continent's speech space of datasets and technologies, leading to a new comprehensive benchmark SimbaBench for downstream African speech tasks. Using SimbaBench, we introduce the Simba family of models, achieving state-of-the-art performance across multiple African languages and speech tasks. Our benchmark analysis reveals critical patterns in resource availability, while our model evaluation demonstrates how dataset quality, domain diversity, and language family relationships influence performance across languages. Our work highlights the need for expanded speech technology resources that better reflect Africa's linguistic diversity and provides a solid foundation for future research and development efforts toward more inclusive speech technologies.

📄 PDF Abstract BibTeX arXiv:2505.18436

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

1000 African Voices: Advancing inclusive multi-speaker multi-accent speech synthesis

2024-06-17 · Sewade Ogun, Abraham T. Owodunni, Tobi Olatunji, Eniola Alese 외

Recent advances in speech synthesis have enabled many useful applications like audio directions in Google Maps, screen readers, and automated content generation on platforms like TikTok. However, these systems are mostly…

DiversitySpeech Synthesis

Mapping the Artificial Intelligence Divide in Africa: Infrastructure, Accessibility and Capacity

2026-06-16 · Abayomi O. Agbeyangi, Jose M. Lukose arxiv

Artificial Intelligence (AI) has the potential to be transformative for development, but Africa is currently facing a fragmented and challenging "AI divide". This paper provides an empirical analysis of the current state…

Synthetic Voice Data for Automatic Speech Recognition in African Languages

2025-07-23 · Brian DeRenzi, Anna Dixon, Mohamed Aymane Farhi, Christian Resch arxiv

Speech technology remains out of reach for most of the over 2300 languages in Africa. We present the first systematic assessment of large-scale synthetic voice corpora for African ASR. We apply a three-step process: LLM-…

Speech Recognition

The NaijaVoices Dataset: Cultivating Large-Scale, High-Quality, Culturally-Rich Speech Data for African Languages

2025-05-26 · Chris Emezue, The NaijaVoices Community, Busayo Awobade, Abraham Owodunni 외

The development of high-performing, robust, and reliable speech technologies depends on large, high-quality datasets. However, African languages -- including our focus, Igbo, Hausa, and Yoruba -- remain under-represented…

Automatic Speech RecognitionDiversityspeech-recognitionSpeech Recognition

AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR

2025-11-18 · Gabrial Zencha Ashungafac, Mardhiyah Sanni, Busayo Awobade, Alex Gichamba 외 arxiv

Recent advances in speech-enabled AI, including Google's NotebookLM and OpenAI's speech-to-speech API, are driving widespread interest in voice interfaces globally. Despite this momentum, there exists no publicly availab…

Speech Recognition