Aggression Identification Using Deep Learning and Data Augmentation
Social media platforms allow users to share and discuss their opinions online. However, a minority of user posts is aggressive, thereby hinders respectful discussion, and {---} at an extreme level {---} is liable to prosecution. The automatic identification of such harmful posts is important, because it can support the costly manual moderation of online discussions. Further, the automation allows unprecedented analyses of discussion datasets that contain millions of posts. This system description paper presents our submission to the First Shared Task on Aggression Identification. We propose to augment the provided dataset to increase the number of labeled comments from 15,000 to 60,000. Thereby, we introduce linguistic variety into the dataset. As a consequence of the larger amount of training data, we are able to train a special deep neural net, which generalizes especially well to unseen data. To further boost the performance, we combine this neural net with three logistic regression classifiers trained on character and word n-grams, and hand-picked syntactic features. This ensemble is more robust than the individual single models. Our team named {``}Julian{''} achieves an F1-score of 60{\%} on both English datasets, 63{\%} on the Hindi Facebook dataset, and 38{\%} on the Hindi Twitter dataset.
Code (1)
Tasks
Aggression IdentificationData AugmentationDeep LearningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Evaluating Aggression Identification in Social Media
In this paper, we present the report and findings of the Shared Task on Aggression and Gendered Aggression Identification organised as part of the Second Workshop on Trolling, Aggression and Cyberbullying (TRAC - 2) at L…
Aggression IdentificationIRIT at TRAC 2020
This paper describes the participation of the IRIT team in the TRAC (Trolling, Aggression and Cyberbullying) 2020 shared task (Bhattacharya et al., 2020) on Aggression Identification and more precisely to the shared task…
Aggression IdentificationLanguage ModelingLanguage ModellingMisogynistic Aggression IdentificationBias, Threat and Aggression Identification Using Machine Learning Techniques on Multilingual Comments
In this paper, we presented our team "IIITRanchi” for the Trolling, Aggression and Cyberbullying (TRAC-3) 2022 shared tasks. Aggression and its different forms on social media and other platforms had tremendous growth on…
Aggression IdentificationPositionThe Role of Computational Stylometry in Identifying (Misogynistic) Aggression in English Social Media Texts
In this paper, we describe UniOr{\_}ExpSys team participation in TRAC-2 (Trolling, Aggression and Cyberbullying) shared task, a workshop organized as part of LREC 2020. TRAC-2 shared task is organized in two sub-tasks: A…
Aggression IdentificationMisogynistic Aggression IdentificationAggression and Misogyny Detection using BERT: A Multi-Task Approach
In recent times, the focus of the NLP community has increased towards offensive language, aggression, and hate-speech detection.This paper presents our system for TRAC-2 shared task on {``}Aggression Identification{''} (…
Abusive LanguageAggression IdentificationHate Speech DetectionMisogynistic Aggression Identification+1