Text Similarity Using Word Embeddings to Classify Misinformation
Fake news is a growing problem in the last years, especially during elections. It's hard work to identify what is true and what is false among all the user generated content that circulates every day. Technology can help with that work and optimize the fact-checking process. In this work, we address the challenge of finding similar content in order to be able to suggest to a fact-checker articles that could have been verified before and thus avoid that the same information is verified more than once. This is especially important in collaborative approaches to fact-checking where members of large teams will not know what content others have already fact-checked.
Code (0)
등록된 구현이 없습니다.
Tasks
ArticlesFact CheckingMisinformationtext similarityWord EmbeddingsSimilar Papers 제목 키워드 기반
Assessing Polyseme Sense Similarity through Co-predication Acceptability and Contextualised Embedding Distance
Co-predication is one of the most frequently used linguistic tests to tell apart shifts in polysemic sense from changes in homonymic meaning. It is increasingly coming under criticism as evidence is accumulating that it …
Word EmbeddingsEating Garlic Prevents COVID-19 Infection: Detecting Misinformation on the Arabic Content of Twitter
The rapid growth of social media content during the current pandemic provides useful tools for disseminating information which has also become a root for misinformation. Therefore, there is an urgent need for fact-checki…
Fact CheckingMisinformationWord EmbeddingsUtility of General and Specific Word Embeddings for Classifying Translational Stages of Research
Conventional text classification models make a bag-of-words assumption reducing text into word occurrence counts per document. Recent algorithms such as word2vec are capable of learning semantic meaning and similarity be…
ArticlesClassificationGeneral Classificationtext-classification+2A Weakly-Supervised Iterative Graph-Based Approach to Retrieve COVID-19 Misinformation Topics
The COVID-19 pandemic has been accompanied by an `infodemic' -- of accurate and inaccurate health information across social media. Detecting misinformation amidst dynamically changing information landscape is challenging…
MisconceptionsMisinformationClassification and Clustering of Arguments with Contextualized Word Embeddings
We experiment with two recent contextualized word embedding methods (ELMo and BERT) in the context of open-domain argument search. For the first time, we show how to leverage the power of contextualized word embeddings t…
Argument MiningClassificationClusteringGeneral Classification+1