paper-with-me

홈 › Papers

Investigation of Multilingual Neural Machine Translation for Indian Languages

2022-10-01 · WAT 2022 10 · Sahinur Rahman Laskar, Riyanka Manna, Partha Pakray, Sivaji Bandyopadhyay

In the domain of natural language processing, machine translation is a well-defined task where one natural language is automatically translated to another natural language. The deep learning-based approach of machine translation, known as neural machine translation attains remarkable translational performance. However, it requires a sufficient amount of training data which is a critical issue for low-resource pair translation. To handle the data scarcity problem, the multilingual concept has been investigated in neural machine translation in different settings like many-to-one and one-to-many translation. WAT2022 (Workshop on Asian Translation 2022) organizes (hosted by the COLING 2022) Indic tasks: English-to-Indic and Indic-to-English translation tasks where we have participated as a team named CNLP-NITS-PP. Herein, we have investigated a transliteration-based approach, where Indic languages are transliterated into English script and shared sub-word level vocabulary during the training phase. We have attained BLEU scores of 2.0 (English-to-Bengali), 1.10 (English-to-Assamese), 4.50 (Bengali-to-English), and 3.50 (Assamese-to-English) translation, respectively.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslationTransliteration

Similar Papers 제목 키워드 기반

Exploring Pair-Wise NMT for Indian Languages

2020-12-10 · ICON 2020 12 · Kartheek Akella, Sai Himal Allu, Sridhar Suresh Ragupathi, Aman Singhal 외

In this paper, we address the task of improving pair-wise machine translation for specific low resource Indian languages. Multilingual NMT models have demonstrated a reasonable amount of effectiveness on resource-poor la…

Machine TranslationNMTTranslation

Machine Translation Advancements of Low-Resource Indian Languages by Transfer Learning

2024-09-24 · Bin Wei, Jiawei Zhen, Zongyao Li, Zhanglin Wu 외

This paper introduces the submission by Huawei Translation Center (HW-TSC) to the WMT24 Indian Languages Machine Translation (MT) Shared Task. To develop a reliable machine translation system for low-resource Indian lang…

Machine TranslationTransfer LearningTranslation

A Large-scale Evaluation of Neural Machine Transliteration for Indic Languages

2021-04-01 · EACL 2021 2 · Anoop Kunchukuttan, Siddharth Jain, Rahul Kejriwal

We take up the task of large-scale evaluation of neural machine transliteration between English and Indic languages, with a focus on multilingual transliteration to utilize orthographic similarity between Indian language…

TranslationTransliteration

Fine-tuning Pre-trained Named Entity Recognition Models For Indian Languages

2024-05-08 · Sankalp Bahad, Pruthwik Mishra, Karunesh Arora, Rakesh Chandra Balabantaray 외

Named Entity Recognition (NER) is a useful component in Natural Language Processing (NLP) applications. It is used in various tasks such as Machine Translation, Summarization, Information Retrieval, and Question-Answerin…

Information RetrievalMachine TranslationMultilingual Named Entity Recognitionnamed-entity-recognition+5

CorIL: Towards Enriching Indian Language to Indian Language Parallel Corpora and Machine Translation Systems

2025-09-24 · Soham Bhattacharjee, Mukund K Roy, Yathish Poojary, Bhargav Dave 외 arxiv

India's linguistic landscape is one of the most diverse in the world, comprising over 120 major languages and approximately 1,600 additional languages, with 22 officially recognized as scheduled languages in the Indian C…

Machine TranslationTransfer LearningDomain Adaptation