paper-with-me

Papers

Quantum Criticism: A Tagged News Corpus Analysed for Sentiment and Named Entities

2020-06-05 · Ashwini Badgujar, Sheng Chen, Andrew Wang, Kai Yu, Paul Intrevado, David Guy Brizan

In this research, we continuously collect data from the RSS feeds of traditional news sources. We apply several pre-trained implementations of named entity recognition (NER) tools, quantifying the success of each implementation. We also perform sentiment analysis of each news article at the document, paragraph and sentence level, with the goal of creating a corpus of tagged news articles that is made available to the public through a web interface. Finally, we show how the data in this corpus could be used to identify bias in news reporting.

📄 PDF Abstract BibTeX arXiv:2006.05267

Code (0)

등록된 구현이 없습니다.

Tasks

Articlesnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NERSentenceSentiment Analysis

Similar Papers 제목 키워드 기반

CroSentiNews 2.0: A Sentence-Level News Sentiment Corpus

2023-05-14 · Gaurish Thakkar, Nives Mikelic Preradović, Marko Tadić

This article presents a sentence-level sentiment dataset for the Croatian news domain. In addition to the 3K annotated texts already present, our dataset contains 14.5K annotated sentence occurrences that have been tagge…

Sentence

SemEval-2019 Task 4: Hyperpartisan News Detection

2019-06-01 · SEMEVAL 2019 6 · Johannes Kiesel, Maria Mestre, Rishabh Shukla, Emmanuel Vincent 외

Hyperpartisan news is news that takes an extreme left-wing or right-wing standpoint. If one is able to reliably compute this meta information, news articles may be automatically tagged, this way encouraging or discouragi…

ArticlesOpen-Ended Question Answeringvalid

Cross-Register Projection for Headline Part of Speech Tagging

2021-09-15 · EMNLP 2021 11 · Adrian Benton, Hanyang Li, Igor Malioutov

Part of speech (POS) tagging is a familiar NLP task. State of the art taggers routinely achieve token-level accuracies of over 97% on news body text, evidence that the problem is well understood. However, the register of…

Open Information ExtractionPart-Of-Speech TaggingPOSPOS Tagging+2

HunOr: A Hungarian---Russian Parallel Corpus

2012-05-01 · LREC 2012 5 · Martina Katalin Szab{\'o}, Veronika Vincze, Istv{\'a}n Nagy T.

In this paper, we present HunOr, the first multi-domain Hungarian―Russian parallel corpus. Some of the corpus texts have been manually aligned and split into sentences, besides, named entities also have been annotated …

Cross-Lingual Information RetrievalInformation RetrievalMachine TranslationNamed Entity Recognition (NER)+4

A Language Model for Grammatical Error Correction in L2 Russian

2023-07-04 · Nikita Remnev, Sergei Obiedkov, Ekaterina Rakhilina, Ivan Smirnov 외

Grammatical error correction is one of the fundamental tasks in Natural Language Processing. For the Russian language, most of the spellcheckers available correct typos and other simple errors with high accuracy, but oft…

Grammatical Error CorrectionLanguage ModelingLanguage Modelling