paper-with-me

Papers

Automatic Syllabification for Manipuri language

2016-12-01 · COLING 2016 12 · Loitongbam Gyanendro Singh, Lenin Laitonjam, Sanasam Ranbir Singh

Development of hand crafted rule for syllabifying words of a language is an expensive task. This paper proposes several data-driven methods for automatic syllabification of words written in Manipuri language. Manipuri is one of the scheduled Indian languages. First, we propose a language-independent rule-based approach formulated using entropy based phonotactic segmentation. Second, we project the syllabification problem as a sequence labeling problem and investigate its effect using various sequence labeling approaches. Third, we combine the effect of sequence labeling and rule-based method and investigate the performance of the hybrid approach. From various experimental observations, it is evident that the proposed methods outperform the baseline rule-based method. The entropy based phonotactic segmentation provides a word accuracy of 96{\%}, CRF (sequence labeling approach) provides 97{\%} and hybrid approach provides 98{\%} word accuracy.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech Recognition (ASR)SegmentationSpeech RecognitionSpeech SynthesisText-To-Speech Synthesis

Methods 이 논문이 사용한 방법론

CRF Conditional Random Fields or CRFs are a type of probabilistic graph model that take neighboring sample context into account for tasks like classification. Prediction is…

Similar Papers 제목 키워드 기반

Language-Agnostic Syllabification with Neural Sequence Labeling

2019-09-29 · Jacob Krantz, Maxwell Dulin, Paul De Palma

The identification of syllables within phonetic sequences is known as syllabification. This task is thought to play an important role in natural language understanding, speech production, and the development of speech re…

Chunkingnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+7

Analysis of Manipuri Tones in ManiTo: A Tonal Contrast Database

2021-12-01 · ICON 2021 12 · Thiyam Susma Devi, Pradip K. Das

Manipuri is a low-resource, tonal language spoken predominantly in Manipur, a northeastern state of India. It has two tones - level and falling tones. For an acceptable Automatic Speech Recognition (ASR) system, integrat…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

An Experiment on Speech-to-Text Translation Systems for Manipuri to English on Low Resource Setting

2021-12-01 · ICON 2021 12 · Loitongbam Sanayai Meetei, Laishram Rahul, Alok Singh, Salam Michael Singh 외

In this paper, we report the experimental findings of building Speech-to-Text translation systems for Manipuri-English on low resource setting which is first of its kind in this language pair. For this purpose, a new dat…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine Translationspeech-recognition+4

Spell-Checking based on Syllabification and Character-level Graphs for a Peruvian Agglutinative Language

2017-09-01 · WS 2017 9 · Carlo Alva, Arturo Oncevay

There are several native languages in Peru which are mostly agglutinative. These languages are transmitted from generation to generation mainly in oral form, causing different forms of writing across different communitie…

EM Corpus: a comparable corpus for a less-resourced language pair Manipuri-English

2021-09-01 · RANLP (BUCC) 2021 9 · Rudali Huidrom, Yves Lepage, Khogendra Khomdram

In this paper, we introduce a sentence-level comparable text corpus crawled and created for the less-resourced language pair, Manipuri(mni) and English (eng). Our monolingual corpora comprise 1.88 million Manipuri senten…

Sentence