paper-with-me

Papers

Natural Language Processing for Tigrinya: Current State and Future Directions

2025-07-23 · Fitsum Gaim, Jong C. Park arxiv

Despite being spoken by millions of people, Tigrinya remains severely underrepresented in Natural Language Processing (NLP) research. This work presents a comprehensive survey of NLP research for Tigrinya, analyzing over 50 studies from 2011 to 2025. We systematically review the current state of computational resources, models, and applications across fifteen downstream tasks, including morphological processing, part-of-speech tagging, named entity recognition, machine translation, question-answering, speech recognition, and synthesis. Our analysis reveals a clear trajectory from foundational, rule-based systems to modern neural architectures, with progress consistently driven by milestones in resource creation. We identify key challenges rooted in Tigrinya's morphological properties and resource scarcity, and highlight promising research directions, including morphology-aware modeling, cross-lingual transfer, and community-centered resource development. This work serves both as a reference for researchers and as a roadmap for advancing Tigrinya NLP. An anthology of surveyed studies and resources is publicly available.

📄 PDF Abstract BibTeX arXiv:2507.17974

Code (0)

등록된 구현이 없습니다.

Tasks

Part-Of-Speech TaggingCross-Lingual TransferMachine TranslationSpeech Recognition

Similar Papers 제목 키워드 기반

Natural Language Processing in Ethiopian Languages: Current State, Challenges, and Opportunities

2023-03-25 · Atnafu Lambebo Tonja, Tadesse Destaw Belay, Israel Abebe Azime, Abinew Ali Ayele 외

This survey delves into the current state of natural language processing (NLP) for four Ethiopian languages: Amharic, Afaan Oromo, Tigrinya, and Wolaytta. Through this paper, we identify key challenges and opportunities …

Transferring Monolingual Model to Low-Resource Language: The Case of Tigrinya

2020-06-13 · Abrhalei Tela, Abraham Woubie, Ville Hautamaki

In recent years, transformer models have achieved great success in natural language processing (NLP) tasks. Most of the current state-of-the-art NLP results are achieved by using monolingual transformer models, where the…

Language ModelingLanguage ModellingSentiment AnalysisTransfer Learning

EthioLLM: Multilingual Large Language Models for Ethiopian Languages with Task Evaluation

2024-03-20 · Atnafu Lambebo Tonja, Israel Abebe Azime, Tadesse Destaw Belay, Mesay Gemeda Yigezu 외

Large language models (LLMs) have gained popularity recently due to their outstanding performance in various downstream Natural Language Processing (NLP) tasks. However, low-resource languages are still lagging behind cu…

Diversity

Tigrinya Number Verbalization: Rules, Algorithm, and Implementation

2026-01-06 · Fitsum Gaim, Issayas Tesfamariam arxiv

We present a systematic formalization of Tigrinya cardinal and ordinal number verbalization, addressing a gap in computational resources for the language. This work documents the canonical rules governing the expression …

Speech Synthesis

Design of a Tigrinya Language Speech Corpus for Speech Recognition

2018-08-01 · COLING 2018 8 · Hafte Abera, Sebsibe H/mariam

In this paper, we describe the first Tigrinya Languages speech corpora designed and development for speech recognition purposes. Tigrinya, often written as Tigrigna (ትግርኛ) /tɪˈɡrinjə/ belongs to the Semitic branch of the…

speech-recognitionSpeech Recognition