paper-with-me

Papers

Tigrinya Number Verbalization: Rules, Algorithm, and Implementation

2026-01-06 · Fitsum Gaim, Issayas Tesfamariam arxiv

We present a systematic formalization of Tigrinya cardinal and ordinal number verbalization, addressing a gap in computational resources for the language. This work documents the canonical rules governing the expression of numerical values in spoken Tigrinya, including the conjunction system, scale words, and special cases for dates, times, and currency. We provide a formal algorithm for number-to-word conversion and release an open-source implementation. Evaluation of frontier large language models (LLMs) reveals significant gaps in their ability to accurately verbalize Tigrinya numbers, underscoring the need for explicit rule documentation. This work serves language modeling, speech synthesis, and accessibility applications targeting Tigrinya-speaking communities.

📄 PDF Abstract BibTeX arXiv:2601.03403

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Synthesis

Similar Papers 제목 키워드 기반

Tigrinya Automatic Speech recognition with Morpheme based recognition units

2020-07-01 · WS 2020 7 · Hafte Abera, sebsibe hailemariam

The Tigrinya language is agglutinative and has a large number of inflected and derived forms of words. Therefore a Tigrinya large vocabulary continuous speech recognition system often has a large number of different unit…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+3

Design of a Tigrinya Language Speech Corpus for Speech Recognition

2018-08-01 · COLING 2018 8 · Hafte Abera, Sebsibe H/mariam

In this paper, we describe the first Tigrinya Languages speech corpora designed and development for speech recognition purposes. Tigrinya, often written as Tigrigna (ትግርኛ) /tɪˈɡrinjə/ belongs to the Semitic branch of the…

speech-recognitionSpeech Recognition

Toward verbalizing ontologies in isiZulu

2014-06-07 · C. Maria Keet, Langa Khumalo

IsiZulu is one of the eleven official languages of South Africa and roughly half the population can speak it. It is the first (home) language for over 10 million people in South Africa. Only a few computational resources…

Machine TranslationSentenceText GenerationTranslation

Natural Language Processing for Tigrinya: Current State and Future Directions

2025-07-23 · Fitsum Gaim, Jong C. Park arxiv

Despite being spoken by millions of people, Tigrinya remains severely underrepresented in Natural Language Processing (NLP) research. This work presents a comprehensive survey of NLP research for Tigrinya, analyzing over…

Part-Of-Speech TaggingCross-Lingual TransferMachine TranslationSpeech Recognition

Tigrinya Neural Machine Translation with Transfer Learning for Humanitarian Response

2020-03-09 · Alp Öktem, Mirko Plitt, Grace Tang

We report our experiments in building a domain-specific Tigrinya-to-English neural machine translation system. We use transfer learning from other Ge'ez script languages and report an improvement of 1.3 BLEU points over …

HumanitarianMachine TranslationTransfer LearningTranslation