paper-with-me

Papers

Detecting Hate Speech and Offensive Language on Twitter using Machine Learning: An N-gram and TFIDF based Approach

2018-09-23 · Aditya Gaydhani, Vikrant Doma, Shrikant Kendre, Laxmi Bhagwat

Toxic online content has become a major issue in today's world due to an exponential increase in the use of internet by people of different cultures and educational background. Differentiating hate speech and offensive language is a key challenge in automatic detection of toxic text content. In this paper, we propose an approach to automatically classify tweets on Twitter into three classes: hateful, offensive and clean. Using Twitter dataset, we perform experiments considering n-grams as features and passing their term frequency-inverse document frequency (TFIDF) values to multiple machine learning models. We perform comparative analysis of the models considering several values of n in n-grams and TFIDF normalization methods. After tuning the model giving the best results, we achieve 95.6% accuracy upon evaluating it on test data. We also create a module which serves as an intermediate between user and Twitter.

📄 PDF Abstract BibTeX arXiv:1809.08651

Code (0)

등록된 구현이 없습니다.

Tasks

Hate Speech Detection

Methods 이 논문이 사용한 방법론

SVM A Support Vector Machine, or SVM, is a non-parametric supervised learning model. For non-linear classification and regression, they utilise the kernel trick to map inputs…

Similar Papers 제목 키워드 기반

Emojis as Anchors to Detect Arabic Offensive Language and Hate Speech

2022-01-18 · Hamdy Mubarak, Sabit Hassan, Shammur Absar Chowdhury

We introduce a generic, language-independent method to collect a large percentage of offensive and hate tweets regardless of their topics or genres. We harness the extralinguistic information embedded in the emojis to co…

Cultural Vocal Bursts Intensity Prediction

Arabic Offensive Language on Twitter: Analysis and Experiments

2020-04-05 · EACL (WANLP) 2021 4 · Hamdy Mubarak, Ammar Rashed, Kareem Darwish, Younes Samih 외

Detecting offensive language on Twitter has many applications ranging from detecting/predicting bullying to measuring polarization. In this paper, we focus on building a large Arabic offensive tweet dataset. We introduce…

aiXplain at Arabic Hate Speech 2022: An Ensemble Based Approach to Detecting Offensive Tweets

2022-06-01 · OSACT (LREC) 2022 6 · Salaheddin Alzubi, Thiago castro Ferreira, Lucas Pavanelli, Mohamed Al-Badrashiny

Abusive speech on online platforms has a detrimental effect on users’ mental health. This warrants the need for innovative solutions that automatically moderate content, especially on online platforms such as Twitter whe…

AlexU-AIC at Arabic Hate Speech 2022: Contrast to Classify

2022-07-18 · OSACT (LREC) 2022 6 · Ahmad Shapiro, Ayman Khalafallah, Marwan Torki

Online presence on social media platforms such as Facebook and Twitter has become a daily habit for internet users. Despite the vast amount of services the platforms offer for their users, users suffer from cyber-bullyin…

Contrastive LearningMulti-Task Learning

Explainable and High-Performance Hate and Offensive Speech Detection

2022-06-26 · Marzieh Babaeianjelodar, Gurram Poorna Prudhvi, Stephen Lorenz, Keyu Chen 외

The spread of information through social media platforms can create environments possibly hostile to vulnerable communities and silence certain groups in society. To mitigate such instances, several models have been deve…

Hate Speech DetectionVocal Bursts Intensity Prediction