paper-with-me

홈 › Papers

Arabic Corpora for Credibility Analysis

2016-05-01 · LREC 2016 5 · Ayman Al Zaatari, Rim El Ballouli, Shady ELbassouni, Wassim El-Hajj, Hazem Hajj, Khaled Shaban, Nizar Habash, Emad Yahya

A significant portion of data generated on blogging and microblogging websites is non-credible as shown in many recent studies. To filter out such non-credible information, machine learning can be deployed to build automatic credibility classifiers. However, as in the case with most supervised machine learning approaches, a sufficiently large and accurate training data must be available. In this paper, we focus on building a public Arabic corpus of blogs and microblogs that can be used for credibility classification. We focus on Arabic due to the recent popularity of blogs and microblogs in the Arab World and due to the lack of any such public corpora in Arabic. We discuss our data acquisition approach and annotation process, provide rigid analysis on the annotated data and finally report some results on the effectiveness of our data for credibility classification.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningGeneral Classification

Similar Papers 제목 키워드 기반

CAT: Credibility Analysis of Arabic Content on Twitter

2017-04-01 · WS 2017 4 · Rim El Ballouli, Wassim El-Hajj, Gh, Ahmad our 외

Data generated on Twitter has become a rich source for various data mining tasks. Those data analysis tasks that are dependent on the tweet semantics, such as sentiment analysis, emotion mining, and rumor detection among…

Emotion RecognitionOpinion MiningSentiment Analysis

Assessing Arabic Weblog Credibility via Deep Co-learning

2019-08-01 · WS 2019 8 · Chadi Helwe, Shady Elbassuoni, Ayman Al Zaatari, Wassim El-Hajj

Assessing the credibility of online content has garnered a lot of attention lately. We focus on one such type of online content, namely weblogs or blogs for short. Some recent work attempted the task of automatically ass…

BIG-bench Machine Learning

Arabic Fake News Detection Based on Deep Contextualized Embedding Models

2022-05-06 · Ali Bou Nassif, Ashraf Elnagar, Omar Elgendy, Yaman Afadar

Social media is becoming a source of news for many people due to its ease and freedom of use. As a result, fake news has been spreading quickly and easily regardless of its credibility, especially in the last decade. Fak…

Fake News Detection

ThatiAR: Subjectivity Detection in Arabic News Sentences

2024-06-08 · Reem Suwaileh, Maram Hasanain, Fatema Hubail, Wajdi Zaghouani 외

Detecting subjectivity in news sentences is crucial for identifying media bias, enhancing credibility, and combating misinformation by flagging opinion-based content. It provides insights into public sentiment, empowers …

In-Context LearningMisinformation

AraWEAT: Multidimensional Analysis of Biases in Arabic Word Embeddings

2020-11-03 · COLING (WANLP) 2020 12 · Anne Lauscher, Rafik Takieddin, Simone Paolo Ponzetto, Goran Glavaš

Recent work has shown that distributional word vector spaces often encode human biases like sexism or racism. In this work, we conduct an extensive analysis of biases in Arabic word embeddings by applying a range of rece…

Word Embeddings