paper-with-me

Papers

Lexical Resources to Enrich English Malayalam Machine Translation

2016-05-01 · LREC 2016 5 · Sreelekha. S, Pushpak Bhattacharyya

In this paper we present our work on the usage of lexical resources for the Machine Translation English and Malayalam. We describe a comparative performance between different Statistical Machine Translation (SMT) systems on top of phrase based SMT system as baseline. We explore different ways of utilizing lexical resources to improve the quality of English Malayalam statistical machine translation. In order to enrich the training corpus we have augmented the lexical resources in two ways (a) additional vocabulary and (b) inflected verbal forms. Lexical resources include IndoWordnet semantic relation set, lexical words and verb phrases etc. We have described case studies, evaluations and have given detailed error analysis for both Malayalam to English and English to Malayalam machine translation systems. We observed significant improvement in evaluations of translation quality. Lexical resources do help uplift performance when parallel corpora are scanty.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Is this Enough?-Evaluation of Malayalam Wordnet

2021-04-01 · EACL (DravidianLangTech) 2021 4 · Nandu Chandran Nair, Maria-chiara Giangregorio, Fausto Giunchiglia

Quality of a product is the degree to which a product meets the customer’s expectation, which must also be valid for the case of lexical semantic resources. Conducting a periodic evaluation of resources is essential to e…

valid

A case study on English-Malayalam Machine Translation

2017-02-27 · Sreelekha. S, Pushpak Bhattacharyya

In this paper we present our work on a case study on Statistical Machine Translation (SMT) and Rule based machine translation (RBMT) for translation from English to Malayalam and Malayalam to English. One of the motivati…

Machine TranslationTranslation

A Sentiment Analysis Dataset for Code-Mixed Malayalam-English

2020-05-30 · LREC 2020 5 · Bharathi Raja Chakravarthi, Navya Jose, Shardul Suryawanshi, Elizabeth Sherly 외

There is an increasing demand for sentiment analysis of text from social media which are mostly code-mixed. Systems trained on monolingual data fail for code-mixed data due to the complexity of mixing at different levels…

Sentiment Analysis

Neural Machine Translation for Malayalam Paraphrase Generation

2024-01-31 · Christeena Varghese, Sergey Koshelev, Ivan P. Yamshchikov

This study explores four methods of generating paraphrases in Malayalam, utilizing resources available for English paraphrasing and pre-trained Neural Machine Translation (NMT) models. We evaluate the resulting paraphras…

Machine TranslationNMTParaphrase GenerationTranslation

CollFrEn: Rich Bilingual English–French Collocation Resource

2020-12-01 · COLING (MWE) 2020 12 · Beatriz Fisas, Luis Espinosa Anke, Joan Codina-Filbá, Leo Wanner

Collocations in the sense of idiosyncratic lexical co-occurrences of two syntactically bound words traditionally pose a challenge to language learners and many Natural Language Processing (NLP) applications alike. Reliab…

Machine TranslationRelation ClassificationText GenerationTranslation+1