French Tweet Corpus for Automatic Stance Detection
The automatic stance detection task consists in determining the attitude expressed in a text toward a target (text, claim, or entity). This is a typical intermediate task for the fake news detection or analysis, which is a considerably widespread and a particularly difficult issue to overcome. This work aims at the creation of a human-annotated corpus for the automatic stance detection of tweets written in French. It exploits a corpus of tweets collected during July and August 2018. To the best of our knowledge, this is the first freely available stance annotated tweet corpus in the French language. The four classes broadly adopted by the community were chosen for the annotation: support, deny, query, and comment with the addition of the ignore class. This paper presents the corpus along with the tools used to build it, its construction, an analysis of the inter-rater reliability, as well as the challenges and questions that were raised during the building process.
Code (0)
등록된 구현이 없습니다.
Tasks
Fake News DetectionStance DetectionSimilar Papers 제목 키워드 기반
A French Corpus for Event Detection on Twitter
We present Event2018, a corpus annotated for event detection tasks, consisting of 38 million tweets in French (retweets excluded) including more than 130,000 tweets manually annotated by three annotators as related or un…
ArticlesEvent DetectionReprésentations lexicales pour la détection non supervisée d'événements dans un flux de tweets : étude sur des corpus français et anglais
In this work, we evaluate the performance of recent text embeddings for the automatic detection of events in a stream of tweets. We model this task as a dynamic clustering problem.Our experiments are conducted on a publi…
ClusteringSentenceAn Annotated Corpus for Sexism Detection in French Tweets
Social media networks have become a space where users are free to relate their opinions and sentiments which may lead to a large spreading of hatred or abusive messages which have to be moderated. This paper presents the…
TREMoLo-Tweets: A Multi-Label Corpus of French Tweets for Language Register Characterization
The casual, neutral, and formal language registers are highly perceptible in discourse productions. However, they are still poorly studied in Natural Language Processing (NLP), especially outside English, and for new tex…
Automatic Classification of Tweets for Analyzing Communication Behavior of Museums
In this paper, we present a study on tweet classification which aims to define the communication behavior of the 103 French museums that participated in 2014 in the Twitter operation: MuseumWeek. The tweets were automati…
ClassificationGeneral Classification