paper-with-me

Papers

An Exploration of Data Augmentation Techniques for Improving English to Tigrinya Translation

2021-03-31 · Lidia Kidane, Sachin Kumar, Yulia Tsvetkov

It has been shown that the performance of neural machine translation (NMT) drops starkly in low-resource conditions, often requiring large amounts of auxiliary data to achieve competitive results. An effective method of generating auxiliary data is back-translation of target language sentences. In this work, we present a case study of Tigrinya where we investigate several back-translation methods to generate synthetic source sentences. We find that in low-resource conditions, back-translation by pivoting through a higher-resource language related to the target language proves most effective resulting in substantial improvements over baselines.

📄 PDF Abstract BibTeX arXiv:2103.16789

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationMachine TranslationNMTTranslation

Similar Papers 제목 키워드 기반

Low-Resource English-Tigrinya MT: Leveraging Multilingual Models, Custom Tokenizers, and Clean Evaluation Benchmarks

2025-09-24 · Hailay Kidu Teklehaymanot, Gebrearegawi Gidey, Wolfgang Nejdl arxiv

Despite advances in Neural Machine Translation (NMT), low-resource languages like Tigrinya remain underserved due to persistent challenges, including limited corpora, inadequate tokenization strategies, and the lack of s…

Machine TranslationTransfer Learning

Tigrinya Neural Machine Translation with Transfer Learning for Humanitarian Response

2020-03-09 · Alp Öktem, Mirko Plitt, Grace Tang

We report our experiments in building a domain-specific Tigrinya-to-English neural machine translation system. We use transfer learning from other Ge'ez script languages and report an improvement of 1.3 BLEU points over …

HumanitarianMachine TranslationTransfer LearningTranslation

Transferring Monolingual Model to Low-Resource Language: The Case of Tigrinya

2020-06-13 · Abrhalei Tela, Abraham Woubie, Ville Hautamaki

In recent years, transformer models have achieved great success in natural language processing (NLP) tasks. Most of the current state-of-the-art NLP results are achieved by using monolingual transformer models, where the…

Language ModelingLanguage ModellingSentiment AnalysisTransfer Learning

TIGQA:An Expert Annotated Question Answering Dataset in Tigrinya

2024-04-26 · Hailay Teklehaymanot, Dren Fazlija, Niloy Ganguly, Gourab K. Patro 외

The absence of explicitly tailored, accessible annotated datasets for educational purposes presents a notable obstacle for NLP tasks in languages with limited resources.This study initially explores the feasibility of us…

Machine TranslationQuestion AnsweringSentence

Tigrinya Number Verbalization: Rules, Algorithm, and Implementation

2026-01-06 · Fitsum Gaim, Issayas Tesfamariam arxiv

We present a systematic formalization of Tigrinya cardinal and ordinal number verbalization, addressing a gap in computational resources for the language. This work documents the canonical rules governing the expression …

Speech Synthesis