Preparation of Sentiment tagged Parallel Corpus and Testing its effect on Machine Translation
In the current work, we explore the enrichment in the machine translation output when the training parallel corpus is augmented with the introduction of sentiment analysis. The paper discusses the preparation of the same sentiment tagged English-Bengali parallel corpus. The preparation of raw parallel corpus, sentiment analysis of the sentences and the training of a Character Based Neural Machine Translation model using the same has been discussed extensively in this paper. The output of the translation model has been compared with a base-line translation model using automated metrics such as BLEU and TER as well as manually.
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationSentiment AnalysisTranslationSimilar Papers 제목 키워드 기반
TALC-sef A Manually-Revised POS-TAgged Literary Corpus in Serbian, English and French
In this paper, we present a parallel literary corpus for Serbian, English and French, the TALC-sef corpus. The corpus includes a manually-revised pos-tagged reference Serbian corpus of over 150,000 words. The initial obj…
POSPOS TaggingQuantum Criticism: A Tagged News Corpus Analysed for Sentiment and Named Entities
In this research, we continuously collect data from the RSS feeds of traditional news sources. We apply several pre-trained implementations of named entity recognition (NER) tools, quantifying the success of each impleme…
Articlesnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+3A Comparison of Sense-level Sentiment Scores
In this paper, we compare a variety of sense-tagged sentiment resources, including SentiWordNet, ML-Senticon, plWordNet emo and the NTU Multilingual Corpus. The goal is to investigate the quality of the resources and see…
Bootstrapping Method for Developing Part-of-Speech Tagged Corpus in Low Resource Languages Tagset - A Focus on an African Igbo
Most languages, especially in Africa, have fewer or no established part-of-speech (POS) tagged corpus. However, POS tagged corpus is essential for natural language processing (NLP) to support advanced researches such as …
Machine TranslationPOSPOS Taggingspeech-recognition+2CroSentiNews 2.0: A Sentence-Level News Sentiment Corpus
This article presents a sentence-level sentiment dataset for the Croatian news domain. In addition to the 3K annotated texts already present, our dataset contains 14.5K annotated sentence occurrences that have been tagge…
Sentence