paper-with-me

홈 › Papers

Preparation of Sentiment tagged Parallel Corpus and Testing its effect on Machine Translation

2020-07-28 · Sainik Kumar Mahata, Amrita Chandra, Dipankar Das, Sivaji Bandyopadhyay

In the current work, we explore the enrichment in the machine translation output when the training parallel corpus is augmented with the introduction of sentiment analysis. The paper discusses the preparation of the same sentiment tagged English-Bengali parallel corpus. The preparation of raw parallel corpus, sentiment analysis of the sentences and the training of a Character Based Neural Machine Translation model using the same has been discussed extensively in this paper. The output of the translation model has been compared with a base-line translation model using automated metrics such as BLEU and TER as well as manually.

📄 PDF Abstract BibTeX arXiv:2007.14074

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationSentiment AnalysisTranslation

Similar Papers 제목 키워드 기반

TALC-sef A Manually-Revised POS-TAgged Literary Corpus in Serbian, English and French

2014-05-01 · LREC 2014 5 · Antonio Balvet, Dejan Stosic, Aleks Miletic, ra

In this paper, we present a parallel literary corpus for Serbian, English and French, the TALC-sef corpus. The corpus includes a manually-revised pos-tagged reference Serbian corpus of over 150,000 words. The initial obj…

POSPOS Tagging

Quantum Criticism: A Tagged News Corpus Analysed for Sentiment and Named Entities

2020-06-05 · Ashwini Badgujar, Sheng Chen, Andrew Wang, Kai Yu 외

In this research, we continuously collect data from the RSS feeds of traditional news sources. We apply several pre-trained implementations of named entity recognition (NER) tools, quantifying the success of each impleme…

Articlesnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+3

A Comparison of Sense-level Sentiment Scores

2019-07-01 · GWC 2019 7 · Francis Bond, Arkadiusz Janz, Maciej Piasecki

In this paper, we compare a variety of sense-tagged sentiment resources, including SentiWordNet, ML-Senticon, plWordNet emo and the NTU Multilingual Corpus. The goal is to investigate the quality of the resources and see…

Bootstrapping Method for Developing Part-of-Speech Tagged Corpus in Low Resource Languages Tagset - A Focus on an African Igbo

2019-03-12 · Onyenwe Ikechukwu E, Onyedinma Ebele G, Aniegwu Godwin E, Ezeani Ignatius M

Most languages, especially in Africa, have fewer or no established part-of-speech (POS) tagged corpus. However, POS tagged corpus is essential for natural language processing (NLP) to support advanced researches such as …

Machine TranslationPOSPOS Taggingspeech-recognition+2

CroSentiNews 2.0: A Sentence-Level News Sentiment Corpus

2023-05-14 · Gaurish Thakkar, Nives Mikelic Preradović, Marko Tadić

This article presents a sentence-level sentiment dataset for the Croatian news domain. In addition to the 3K annotated texts already present, our dataset contains 14.5K annotated sentence occurrences that have been tagge…

Sentence