paper-with-me

Papers

Multi-Task Learning using AraBert for Offensive Language Detection

2020-05-01 · LREC 2020 5 · Dj, Marc ji, Fady Baly, Wissam Antoun, Hazem Hajj

The use of social media platforms has become more prevalent, which has provided tremendous opportunities for people to connect but has also opened the door for misuse with the spread of hate speech and offensive language. This phenomenon has been driving more and more people to more extreme reactions and online aggression, sometimes causing physical harm to individuals or groups of people. There is a need to control and prevent such misuse of online social media through automatic detection of profane language. The shared task on Offensive Language Detection at the OSACT4 has aimed at achieving state of art profane language detection methods for Arabic social media. Our team {``}BERTologists{''} tackled this problem by leveraging state of the art pretrained Arabic language model, AraBERT, that we augment with the addition of Multi-task learning to enable our model to learn efficiently from little data. Our Multitask AraBERT approach achieved the second place in both subtasks A {\&} B, which shows that the model performs consistently across different tasks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingMulti-Task Learning

Similar Papers 제목 키워드 기반

LISAC FSDM-USMBA Team at SemEval-2020 Task 12: Overcoming AraBERT's pretrain-finetune discrepancy for Arabic offensive language identification

2020-12-01 · SEMEVAL 2020 · Hamza Alami, Said Ouatik El Alaoui, Abdessamad Benlahbib, Noureddine En-nahnahi

AraBERT is an Arabic version of the state-of-the-art Bidirectional Encoder Representations from Transformers (BERT) model. The latter has achieved good performance in a variety of Natural Language Processing (NLP) tasks.…

Language Identification

An Annotated Corpus of Arabic Tweets for Hate Speech Analysis

2025-05-17 · Md. Rafiul Biswas, Wajdi Zaghouani

Identifying hate speech content in the Arabic language is challenging due to the rich quality of dialectal variations. This study introduces a multilabel hate speech dataset in the Arabic language. We have collected 1000…

UPV at the Arabic Hate Speech 2022 Shared Task: Offensive Language and Hate Speech Detection using Transformers and Ensemble Models

2022-06-01 · OSACT (LREC) 2022 6 · Angel Felipe Magnossão de Paula, Paolo Rosso, Imene Bensalem, Wajdi Zaghouani

This paper describes our participation in the shared task Fine-Grained Hate Speech Detection on Arabic Twitter at the 5th Workshop on Open-Source Arabic Corpora and Processing Tools (OSACT). The shared task is divided in…

DecoderHate Speech Detection

AraBERT: Transformer-based Model for Arabic Language Understanding

2020-02-28 · LREC 2020 5 · Wissam Antoun, Fady Baly, Hazem Hajj

The Arabic language is a morphologically rich language with relatively few resources and a less explored syntax compared to English. Given these limitations, Arabic Natural Language Processing (NLP) tasks like Sentiment …

modelnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+4

Quranic Verses Semantic Relatedness Using AraBERT

2021-04-01 · EACL (WANLP) 2021 4 · Abdullah Alsaleh, Eric Atwell, Abdulrahman Altahhan

Bidirectional Encoder Representations from Transformers (BERT) has gained popularity in recent years producing state-of-the-art performances across Natural Language Processing tasks. In this paper, we used AraBERT langua…

Language ModelingLanguage Modelling