paper-with-me

Papers

Crowdsourcing Salient Information from News and Tweets

2016-05-01 · LREC 2016 5 · Oana Inel, Tommaso Caselli, Lora Aroyo

The increasing streams of information pose challenges to both humans and machines. On the one hand, humans need to identify relevant information and consume only the information that lies at their interests. On the other hand, machines need to understand the information that is published in online data streams and generate concise and meaningful overviews. We consider events as prime factors to query for information and generate meaningful context. The focus of this paper is to acquire empirical insights for identifying salience features in tweets and news about a target event, i.e., the event of {``}whaling{''}. We first derive a methodology to identify such features by building up a knowledge space of the event enriched with relevant phrases, sentiments and ranked by their novelty. We applied this methodology on tweets and we have performed preliminary work towards adapting it to news articles. Our results show that crowdsourcing text relevance, sentiments and novelty (1) can be a main step in identifying salient information, and (2) provides a deeper and more precise understanding of the data at hand compared to state-of-the-art approaches.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Articles

Similar Papers 제목 키워드 기반

The NewSoMe Corpus: A Unifying Opinion Annotation Framework across Genres and in Multiple Languages

2014-05-01 · LREC 2014 5 · Roser Saur{\'\i}, Judith Domingo, Toni Badia

We present the NewSoMe (News and Social Media) Corpus, a set of subcorpora with annotations on opinion expressions across genres (news reports, blogs, product reviews and tweets) and covering multiple languages (English,…

Information RetrievalOpinion Mining

CoVERT: A Corpus of Fact-checked Biomedical COVID-19 Tweets

2022-04-26 · LREC 2022 6 · Isabelle Mohr, Amelie Wührl, Roman Klinger

Over the course of the COVID-19 pandemic, large volumes of biomedical information concerning this new disease have been published on social media. Some of this information can pose a real danger to people's health, parti…

Fact CheckingMisinformation

Crowdsourcing a Large Corpus of Clickbait on Twitter

2018-08-01 · COLING 2018 8 · Martin Potthast, Tim Gollub, Kristof Komlossy, Sebastian Schuster 외

Clickbait has become a nuisance on social media. To address the urging task of clickbait detection, we constructed a new corpus of 38,517 annotated Twitter tweets, the Webis Clickbait Corpus 2017. To avoid biases in term…

Clickbait Detection

Linking Tweets with Monolingual and Cross-Lingual News using Transformed Word Embeddings

2017-10-25 · Aditya Mogadala, Dominik Jung, Achim Rettinger

Social media platforms have grown into an important medium to spread information about an event published by the traditional media, such as news articles. Grouping such diverse sources of information that discuss the sam…

ArticlesWord Embeddings

WN-Salience: A Corpus of News Articles with Entity Salience Annotations

2020-05-01 · LREC 2020 5 · Chuan Wu, Evangelos Kanoulas, Maarten de Rijke, Wei Lu

Entities can be found in various text genres, ranging from tweets and web pages to user queries submitted to web search engines. Existing research either considers all entities in the text equally important, or heuristic…

ArticlesEntity Linking