paper-with-me

Papers

Improve Sinhala Speech Recognition Through e2e LF-MMI Model

2021-12-01 · ICON 2021 12 · Buddhi Gamage, Randil Pushpananda, Thilini Nadungodage, Ruwan Weerasinghe

Automatic speech recognition (ASR) has experienced several paradigm shifts over the years from template-based approaches and statistical modeling to the popular GMM-HMM approach and then to deep learning hybrid model DNN-HMM. The latest shift is to end-to-end (e2e) DNN architecture. We present a study to build an e2e ASR system using state-of-the-art deep learning models to verify the applicability of e2e ASR models for the highly inflected and yet low-resource Sinhala language. We experimented on e2e Lattice-Free Maximum Mutual Information (e2e LF-MMI) model with the baseline statistical models with 40 hours of training data to evaluate. We used the same corpus for creating language models and lexicon in our previous study, which resulted in the best accuracy for the Sinhala language. We were able to achieve a Word-error-rate (WER) of 28.55% for Sinhala, only slightly worse than the existing best hybrid model. Our model, however, is more context-independent and faster for Sinhala speech recognition and so more suitable for general purpose speech-to-text translation.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech RecognitionSpeech-to-TextSpeech-to-Text Translation

Similar Papers 제목 키워드 기반

A Systematic Approach to Derive a Refined Speech Corpus for Sinhala

2022-06-01 · LREC 2022 6 · Disura Warusawithana, Nilmani Kulaweera, Lakshan Weerasinghe, Buddhika Karunarathne

Speech Recognition is an active research area where advances of technology have continuously driven the development of research work. However, due to the lack of adequate resources, certain languages such as Sinhala, are…

speech-recognitionSpeech Recognition

From Sinhala to Dhivehi: Cross-Lingual Transfer Learning for Low-Resource Speech Recognition

2026-07-07 · Lukmal Ilyas, Nevidu Jayatilleke arxiv

Dhivehi, the national language of the Maldives, is currently under-resourced for automatic speech recognition (ASR) and other NLP tasks. This study investigates whether cross-lingual transfer learning from Sinhala, a lin…

Cross-Lingual TransferSpeech RecognitionTransfer Learning

A Fuzzy Based Model to Identify Printed Sinhala Characters (ICIAfS14)

2014-12-24 · G. I. Gunarathna, M. A. P. Chamikara, R. G. Ragel

Character recognition techniques for printed documents are widely used for English language. However, the systems that are implemented to recognize Asian languages struggle to increase the accuracy of recognition. Among …

Comprehensive Part-Of-Speech Tag Set and SVM based POS Tagger for Sinhala

2016-12-01 · WS 2016 12 · Fern, S o, areka, Surangika Ranathunga 외

This paper presents a new comprehensive multi-level Part-Of-Speech tag set and a Support Vector Machine based Part-Of-Speech tagger for the Sinhala language. The currently available tag set for Sinhala has two limitation…

POSTAG

Evaluation of Noise Reduction Methods for Sentence Recognition by Sinhala Speaking Listeners

2023-03-31 · Malitha Gunawardhana, Chathuki Navanjana, Dinithi Fernando, Nipuna Upeksha 외

Noise reduction is a crucial aspect of hearing aids, which researchers have been striving to address over the years. However, most existing noise reduction algorithms have primarily been evaluated using English. Consider…

Action DetectionActivity DetectionSentence