Unsupervised corpus--wide claim detection
Automatic claim detection is a fundamental argument mining task that aims to automatically mine claims regarding a topic of consideration. Previous works on mining argumentative content have assumed that a set of relevant documents is given in advance. Here, we present a first corpus{--} wide claim detection framework, that can be directly applied to massive corpora. Using simple and intuitive empirical observations, we derive a claim sentence query by which we are able to directly retrieve sentences in which the prior probability to include topic-relevant claims is greatly enhanced. Next, we employ simple heuristics to rank the sentences, leading to an unsupervised corpus{--}wide claim detection system, with precision that outperforms previously reported results on the task of claim detection given relevant documents and labeled data.
Code (0)
등록된 구현이 없습니다.
Tasks
Argument MiningDecision MakingSentenceSimilar Papers 제목 키워드 기반
Towards Effective Rebuttal: Listening Comprehension using Corpus-Wide Claim Mining
Engaging in a live debate requires, among other things, the ability to effectively rebut arguments claimed by your opponent. In particular, this requires identifying these arguments. Here, we suggest doing so by automati…
ArticlesFrench Tweet Corpus for Automatic Stance Detection
The automatic stance detection task consists in determining the attitude expressed in a text toward a target (text, claim, or entity). This is a typical intermediate task for the fake news detection or analysis, which is…
Fake News DetectionStance DetectionClaim Detection in Biomedical Twitter Posts
Social media contains unfiltered and unique information, which is potentially of great value, but, in the case of misinformation, can also do great harm. With regards to biomedical topics, false information can be partic…
Fact CheckingFake News DetectionMisinformationTransfer LearningIMHO Fine-Tuning Improves Claim Detection
Claims are the central component of an argument. Detecting claims across different domains or data sets can often be challenging due to their varying conceptualization. We propose to alleviate this problem by fine tuning…
Language ModelingLanguage ModellingSentimental LIAR: Extended Corpus and Deep Learning Models for Fake Claim Classification
The rampant integration of social media in our every day lives and culture has given rise to fast and easier access to the flow of information than ever in human history. However, the inherently unsupervised nature of so…
Cultural Vocal Bursts Intensity PredictionEmotion RecognitionFake News DetectionGeneral Classification+2