paper-with-me

홈 › Papers

Automatic Speech Recognition for Humanitarian Applications in Somali

2018-07-23 · Raghav Menon, Astik Biswas, Armin Saeb, John Quinn, Thomas Niesler

We present our first efforts in building an automatic speech recognition system for Somali, an under-resourced language, using 1.57 hrs of annotated speech for acoustic model training. The system is part of an ongoing effort by the United Nations (UN) to implement keyword spotting systems supporting humanitarian relief programmes in parts of Africa where languages are severely under-resourced. We evaluate several types of acoustic model, including recent neural architectures. Language model data augmentation using a combination of recurrent neural networks (RNN) and long short-term memory neural networks (LSTMs) as well as the perturbation of acoustic data are also considered. We find that both types of data augmentation are beneficial to performance, with our best system using a combination of convolutional neural networks (CNNs), time-delay neural networks (TDNNs) and bi-directional long short term memory (BLSTMs) to achieve a word error rate of 53.75%.

📄 PDF Abstract BibTeX arXiv:1807.08669

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationHumanitarianKeyword SpottingLanguage ModelingLanguage Modellingspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Improved low-resource Somali speech recognition by semi-supervised acoustic and language model training

2019-07-06 · Astik Biswas, Raghav Menon, Ewald van der Westhuizen, Thomas Niesler

We present improvements in automatic speech recognition (ASR) for Somali, a currently extremely under-resourced language. This forms part of a continuing United Nations (UN) effort to employ ASR-based keyword spotting sy…

Acoustic ModellingAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Decoder+6

Fast Development of ASR in African Languages using Self Supervised Speech Representation Learning

2021-03-16 · Jama Hussein Mohamud, Lloyd Acquaye Thompson, Aissatou Ndoye, Laurent Besacier

This paper describes the results of an informal collaboration launched during the African Master of Machine Intelligence (AMMI) in June 2020. After a series of lectures and labs on speech data collection using mobile app…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Representation Learningspeech-recognition+2

Automatic Speech Recognition and Topic Identification for Almost-Zero-Resource Languages

2018-02-23 · Matthew Wiesner, Chunxi Liu, Lucas Ondel, Craig Harman 외

Automatic speech recognition (ASR) systems often need to be developed for extremely low-resource languages to serve end-uses such as audio content categorization and search. While universal phone recognition is natural t…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Humanitarianspeech-recognition+1

AfriVoices-KE: A Multilingual Speech Dataset for Kenyan Languages

2026-04-09 · Lilian Wanzare, Cynthia Amol, Ezekiel Maina, Nelson Odhiambo 외 arxiv

AfriVoices-KE is a large-scale multilingual speech dataset comprising approximately 3,000 hours of audio across five Kenyan languages: Dholuo, Kikuyu, Kalenjin, Maasai, and Somali. The dataset includes 750 hours of scrip…

Speech Recognition

Corpora for Cross-Language Information Retrieval in Six Less-Resourced Languages

2020-05-01 · LREC 2020 5 · Ilya Zavorin, Aric Bills, Cassian Corey, Michelle Morrison 외

The Machine Translation for English Retrieval of Information in Any Language (MATERIAL) research program, sponsored by the Intelligence Advanced Research Projects Activity (IARPA), focuses on rapid development of end-to-…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Information RetrievalMachine Translation+4