paper-with-me

홈 › Papers

CAT: Credibility Analysis of Arabic Content on Twitter

2017-04-01 · WS 2017 4 · Rim El Ballouli, Wassim El-Hajj, Gh, Ahmad our, Shady Elbassuoni, Hazem Hajj, Khaled Shaban

Data generated on Twitter has become a rich source for various data mining tasks. Those data analysis tasks that are dependent on the tweet semantics, such as sentiment analysis, emotion mining, and rumor detection among others, suffer considerably if the tweet is not credible, not real, or spam. In this paper, we perform an extensive analysis on credibility of Arabic content on Twitter. We also build a classification model (CAT) to automatically predict the credibility of a given Arabic tweet. Of particular originality is the inclusion of features extracted directly or indirectly from the author{'}s profile and timeline. To train and test CAT, we annotated for credibility a data set of 9,000 Arabic tweets that are topic independent. CAT achieved consistent improvements in predicting the credibility of the tweets when compared to several baselines and when compared to the state-of-the-art approach with an improvement of 21{\%} in weighted average F-measure. We also conducted experiments to highlight the importance of the user-based features as opposed to the content-based features. We conclude our work with a feature reduction experiment that highlights the best indicative features of credibility.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionOpinion MiningSentiment Analysis

Similar Papers 제목 키워드 기반

Adult Content Detection on Arabic Twitter: Analysis and Experiments

2021-04-01 · EACL (WANLP) 2021 4 · Hamdy Mubarak, Sabit Hassan, Ahmed Abdelali

With Twitter being one of the most popular social media platforms in the Arab region, it is not surprising to find accounts that post adult content in Arabic tweets; despite the fact that these platforms dissuade users f…

Assessing Arabic Weblog Credibility via Deep Co-learning

2019-08-01 · WS 2019 8 · Chadi Helwe, Shady Elbassuoni, Ayman Al Zaatari, Wassim El-Hajj

Assessing the credibility of online content has garnered a lot of attention lately. We focus on one such type of online content, namely weblogs or blogs for short. Some recent work attempted the task of automatically ass…

BIG-bench Machine Learning

Hateful People or Hateful Bots? Detection and Characterization of Bots Spreading Religious Hatred in Arabic Social Media

2019-08-01 · Nuha Albadi, Maram Kurdi, Shivakant Mishra

Arabic Twitter space is crawling with bots that fuel political feuds, spread misinformation, and proliferate sectarian rhetoric. While efforts have long existed to analyze and detect English bots, Arabic bot detection an…

Misinformation

ArabGend: Gender Analysis and Inference on Arabic Twitter

2022-03-01 · COLING (WNUT) 2022 10 · Hamdy Mubarak, Shammur Absar Chowdhury, Firoj Alam

Gender analysis of Twitter can reveal important socio-cultural differences between male and female users. There has been a significant effort to analyze and automatically infer gender in the past for most widely spoken l…

Arabic Corpora for Credibility Analysis

2016-05-01 · LREC 2016 5 · Ayman Al Zaatari, Rim El Ballouli, Shady ELbassouni, Wassim El-Hajj 외

A significant portion of data generated on blogging and microblogging websites is non-credible as shown in many recent studies. To filter out such non-credible information, machine learning can be deployed to build autom…

BIG-bench Machine LearningGeneral Classification