IIITT at CASE 2021 Task 1: Leveraging Pretrained Language Models for Multilingual Protest Detection
In a world abounding in constant protests resulting from events like a global pandemic, climate change, religious or political conflicts, there has always been a need to detect events/protests before getting amplified by news media or social media. This paper demonstrates our work on the sentence classification subtask of multilingual protest detection in CASE@ACL-IJCNLP 2021. We approached this task by employing various multilingual pre-trained transformer models to classify if any sentence contains information about an event that has transpired or not. We performed soft voting over the models, achieving the best results among the models, accomplishing a macro F1-Score of 0.8291, 0.7578, and 0.7951 in English, Spanish, and Portuguese, respectively.
Code (1)
Tasks
SentenceSentence ClassificationSimilar Papers 제목 키워드 기반
Attentive fine-tuning of Transformers for Translation of low-resourced languages @LoResMT 2021
This paper reports the Machine Translation (MT) systems submitted by the IIITT team for the English->Marathi and English->Irish language pairs LoResMT 2021 shared task. The task focuses on getting exceptional translation…
Machine TranslationNMTTranslationIIITT@DravidianLangTech-EACL2021: Transfer Learning for Offensive Language Detection in Dravidian Languages
This paper demonstrates our work for the shared task on Offensive Language Identification in Dravidian Languages-EACL 2021. Offensive language detection in the various social media platforms was identified previously. Bu…
DiversityLanguage IdentificationTransfer LearningIIITT@LT-EDI-EACL2021-Hope Speech Detection: There is always Hope in Transformers
In a world filled with serious challenges like climate change, religious and political conflicts, global pandemics, terrorism, and racial discrimination, an internet full of hate speech, abusive and offensive content is …
DiversityHope Speech DetectionCombining Pretrained High-Resource Embeddings and Subword Representations for Low-Resource Languages
The contrast between the need for large amounts of data for current Natural Language Processing (NLP) techniques, and the lack thereof, is accentuated in the case of African languages, most of which are considered low-re…
TranslationWord EmbeddingsUVCE-IIITT@DravidianLangTech-EACL2021: Tamil Troll Meme Classification: You need to Pay more Attention
Tamil is a Dravidian language that is commonly used and spoken in the southern part of Asia. In the era of social media, memes have been a fun moment in the day-to-day life of people. Here, we try to analyze the true mea…
Binary ClassificationGeneral ClassificationMeme Classification