paper-with-me

홈 › Papers

Did you offend me? Classification of Offensive Tweets in Hinglish Language

2018-10-01 · WS 2018 10 · Puneet Mathur, Ramit Sawhney, Meghna Ayyar, Rajiv Shah

The use of code-switched languages (\textit{e.g.}, Hinglish, which is derived by the blending of Hindi with the English language) is getting much popular on Twitter due to their ease of communication in native languages. However, spelling variations and absence of grammar rules introduce ambiguity and make it difficult to understand the text automatically. This paper presents the Multi-Input Multi-Channel Transfer Learning based model (MIMCT) to detect offensive (hate speech or abusive) Hinglish tweets from the proposed Hinglish Offensive Tweet (HOT) dataset using transfer learning coupled with multiple feature inputs. Specifically, it takes multiple primary word embedding along with secondary extracted features as inputs to train a multi-channel CNN-LSTM architecture that has been pre-trained on English tweets through transfer learning. The proposed MIMCT model outperforms the baseline supervised classification models, transfer learning based CNN and LSTM models to establish itself as the state of the art in the unexplored domain of Hinglish offensive text classification.

📄 PDF Abstract BibTeX

Code (1)

pmathur5k10/Hinglish-Offensive-Text-Classification 공식 구현

Tasks

Abuse DetectionGeneral Classificationtext-classificationText ClassificationTransfer Learning

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

Detecting Offensive Tweets in Hindi-English Code-Switched Language

2018-07-01 · WS 2018 7 · Puneet Mathur, Rajiv Shah, Ramit Sawhney, Debanjan Mahata

The exponential rise of social media websites like Twitter, Facebook and Reddit in linguistically diverse geographical regions has led to hybridization of popular native languages with English in an effort to ease commun…

General ClassificationHate Speech DetectionTransfer Learning

QutNocturnal@HASOC'19: CNN for Hate Speech and Offensive Content Identification in Hindi Language

2020-08-28 · Md Abul Bashar, Richi Nayak

We describe our top-team solution to Task 1 for Hindi in the HASOC contest organised by FIRE 2019. The task is to identify hate speech and offensive language in Hindi. More specifically, it is a binary classification pro…

Binary Classification

HateGPT: Unleashing GPT-3.5 Turbo to Combat Hate Speech on X

2024-11-14 · Aniket Deroy, Subhankar Maity

The widespread use of social media platforms like Twitter and Facebook has enabled people of all ages to share their thoughts and experiences, leading to an immense accumulation of user-generated content. However, alongs…

Mind Your Language: Abuse and Offense Detection for Code-Switched Languages

2018-09-23 · Raghav Kapoor, Yaman Kumar, Kshitij Rajput, Rajiv Ratn Shah 외

In multilingual societies like the Indian subcontinent, use of code-switched languages is much popular and convenient for the users. In this paper, we study offense and abuse detection in the code-switched pair of Hindi …

Abuse DetectionGeneral ClassificationTransfer Learning

Code-Mix Sentiment Analysis on Hinglish Tweets

2026-01-08 · Aashi Garg, Aneshya Das, Arshi Arya, Anushka Goyal 외 arxiv

The effectiveness of brand monitoring in India is increasingly challenged by the rise of Hinglish--a hybrid of Hindi and English--used widely in user-generated content on platforms like Twitter. Traditional Natural Langu…

Sentiment Analysis