paper-with-me

Papers

Large vocabulary speech recognition for languages of Africa: multilingual modeling and self-supervised learning

2022-08-05 · Sandy Ritchie, You-Chi Cheng, Mingqing Chen, Rajiv Mathews, Daan van Esch, Bo Li, Khe Chai Sim

Almost none of the 2,000+ languages spoken in Africa have widely available automatic speech recognition systems, and the required data is also only available for a few languages. We have experimented with two techniques which may provide pathways to large vocabulary speech recognition for African languages: multilingual modeling and self-supervised learning. We gathered available open source data and collected data for 15 languages, and trained experimental models using these techniques. Our results show that pooling the small amounts of data available in multilingual end-to-end models, and pre-training on unsupervised data can help improve speech recognition quality for many African languages.

📄 PDF Abstract BibTeX arXiv:2208.03067

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Self-Supervised Learningspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

AfroDigits: A Community-Driven Spoken Digit Dataset for African Languages

2023-03-22 · Chris Chinenye Emezue, Sanchit Gandhi, Lewis Tunstall, Abubakar Abid 외

The advancement of speech technologies has been remarkable, yet its integration with African languages remains limited due to the scarcity of African speech corpora. To address this issue, we present AfroDigits, a minima…

OkwuGbé: End-to-End Speech Recognition for Fon and Igbo

2021-10-16 · ACL ARR October 2021 10 · Anonymous

Language is inherent and compulsory for human communication. Whether expressed in a written or spoken way, it ensures understanding between people of the same and different regions. With the growing awareness and effort …

Machine Translationspeech-recognitionSpeech Recognition

OkwuGbé: End-to-End Speech Recognition for Fon and Igbo

2021-03-13 · Bonaventure F. P. Dossou, Chris C. Emezue

Language is inherent and compulsory for human communication. Whether expressed in a written or spoken way, it ensures understanding between people of the same and different regions. With the growing awareness and effort …

Machine Translationspeech-recognitionSpeech Recognition

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition

2026-05-05 · Busayo Awobade, Gabrial Zencha Ashungafac, Tobi Olatunji arxiv

Recent large language models (LLMs) show strong speech recognition and translation capabilities for high-resource languages. However, African languages remain dramatically underrepresented in benchmarks, limiting their p…

Speech Recognition

The NaijaVoices Dataset: Cultivating Large-Scale, High-Quality, Culturally-Rich Speech Data for African Languages

2025-05-26 · Chris Emezue, The NaijaVoices Community, Busayo Awobade, Abraham Owodunni 외

The development of high-performing, robust, and reliable speech technologies depends on large, high-quality datasets. However, African languages -- including our focus, Igbo, Hausa, and Yoruba -- remain under-represented…

Automatic Speech RecognitionDiversityspeech-recognitionSpeech Recognition