paper-with-me

홈 › Papers

Tigrinya Automatic Speech recognition with Morpheme based recognition units

2020-07-01 · WS 2020 7 · Hafte Abera, sebsibe hailemariam

The Tigrinya language is agglutinative and has a large number of inflected and derived forms of words. Therefore a Tigrinya large vocabulary continuous speech recognition system often has a large number of different units and a high out-of-vocabulary (OOV) rate if a word is used as a recognition unit of a language model (LM) and lexicon. Therefore a morpheme-based approach has often been used and a morpheme is used as the recognition unit to reduce the high OOV rate. This paper presents an automatic speech recognition experiment conducted to see the effect of OOV words on the performance speech recognition system for Tigrinya. We tried to solve the OOV problem by using morphemes as lexicon and language model units. It has been found that the morpheme-based recognition system is better lexical and language modeling units than words. An absolute improvement (in word recognition accuracy) of 3.45 token and 8.36 types has been obtained as a result of using a morph-based vocabulary.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage ModellingMORPHspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Design of a Tigrinya Language Speech Corpus for Speech Recognition

2018-08-01 · COLING 2018 8 · Hafte Abera, Sebsibe H/mariam

In this paper, we describe the first Tigrinya Languages speech corpora designed and development for speech recognition purposes. Tigrinya, often written as Tigrigna (ትግርኛ) /tɪˈɡrinjə/ belongs to the Semitic branch of the…

speech-recognitionSpeech Recognition

Amharic-English Speech Translation in Tourism Domain

2017-09-01 · WS 2017 9 · Michael Melese, Laurent Besacier, Million Meshesha

This paper describes speech translation from Amharic-to-English, particularly Automatic Speech Recognition (ASR) with post-editing feature and Amharic-English Statistical Machine Translation (SMT). ASR experiment is cond…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+4

Natural Language Processing for Tigrinya: Current State and Future Directions

2025-07-23 · Fitsum Gaim, Jong C. Park arxiv

Despite being spoken by millions of people, Tigrinya remains severely underrepresented in Natural Language Processing (NLP) research. This work presents a comprehensive survey of NLP research for Tigrinya, analyzing over…

Part-Of-Speech TaggingCross-Lingual TransferMachine TranslationSpeech Recognition

Speech Recognition for Tigrinya language Using Deep Neural Network Approach

2019-08-01 · WS 2019 8 · Hafte Abera, Sebsibe H/mariam

This work presents a speech recognition model for Tigrinya language .The Deep Neural Network is used to make the recognition model. The Long Short-Term Memory Network (LSTM), which is a special kind of Recurrent Neural N…

speech-recognitionSpeech Recognition

Ethio-ASR: Joint Multilingual Speech Recognition and Language Identification for Ethiopian Languages

2026-03-24 · Badr M. Abdullah, Israel Abebe Azime, Atnafu Lambebo Tonja, Jesujoba O. Alabi 외 arxiv

We present Ethio-ASR, a suite of multilingual CTC-based automatic speech recognition (ASR) models jointly trained on five Ethiopian languages: Amharic, Tigrinya, Oromo, Sidaama, and Wolaytta. These languages belong to th…

Language IdentificationSpeech Recognition