paper-with-me

홈 › Papers

TuPy-E: detecting hate speech in Brazilian Portuguese social media with a novel dataset and comprehensive analysis of models

2023-12-29 · Felipe Oliveira, Victoria Reis, Nelson Ebecken

Social media has become integral to human interaction, providing a platform for communication and expression. However, the rise of hate speech on these platforms poses significant risks to individuals and communities. Detecting and addressing hate speech is particularly challenging in languages like Portuguese due to its rich vocabulary, complex grammar, and regional variations. To address this, we introduce TuPy-E, the largest annotated Portuguese corpus for hate speech detection. TuPy-E leverages an open-source approach, fostering collaboration within the research community. We conduct a detailed analysis using advanced techniques like BERT models, contributing to both academic understanding and practical applications

📄 PDF Abstract BibTeX arXiv:2312.17704

Code (1)

Silly-Machine/TuPyE-Dataset 공식 구현

Tasks

Hate Speech Detection

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Multi-Head Attention 설명 없음
Attention 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Weight Decay 설명 없음
WordPiece 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…

Similar Papers 제목 키워드 기반

HateBR: A Large Expert Annotated Corpus of Brazilian Instagram Comments for Offensive Language and Hate Speech Detection

2021-03-27 · LREC 2022 6 · Francielle Alves Vargas, Isabelle Carvalho, Fabiana Rodrigues de Góes, Fabrício Benevenuto 외

Due to the severity of the social media offensive and hateful comments in Brazil, and the lack of research in Portuguese, this paper provides the first large-scale expert annotated corpus of Brazilian Instagram comments …

BIG-bench Machine LearningBinary ClassificationHate Speech Detection

HateBR: Large expert annotated corpus of Brazilian Instagram comments for abusive language detection

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Due to the severity of the social media abusive comments in Brazil, and the lack of research in Portuguese, this paper provides the first large-scale annotated corpus of Brazilian Instagram comments for hate speech and o…

Abusive LanguageBinary Classification

Toxic Language Detection in Social Media for Brazilian Portuguese: New Dataset and Multilingual Analysis

2020-10-09 · Asian Chapter of the Association for Computational Linguistics 2020 · João A. Leite, Diego F. Silva, Kalina Bontcheva, Carolina Scarton

Hate speech and toxic comments are a common concern of social media platform users. Although these comments are, fortunately, the minority in these platforms, they are still capable of causing harm. Therefore, identifyin…

Hate Speech DetectionMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATIONText Classification

``Haters gonna hate'': challenges for sentiment analysis of Facebook comments in Brazilian Portuguese

2017-09-01 · WS 2017 9 · Juliano D. Antonio, Ana Carolina L. Santin
Sentiment Analysis

MIN_PT: An European Portuguese Lexicon for Minorities Related Terms

2021-08-01 · ACL (WOAH) 2021 8 · Paula Fortuna, Vanessa Cortez, Miguel Sozinho Ramalho, Laura Pérez-Mayos

Hate speech-related lexicons have been proved to be useful for many tasks such as data collection and classification. However, existing Portuguese lexicons do not distinguish between European and Brazilian Portuguese, an…