paper-with-me

Papers

HateBR: Large expert annotated corpus of Brazilian Instagram comments for abusive language detection

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Due to the severity of the social media abusive comments in Brazil, and the lack of research in Portuguese, this paper provides the first large-scale annotated corpus of Brazilian Instagram comments for hate speech and offensive language detection on the web and social media. The HateBR corpus was collected from Brazilian Instagram comments of political personalities and manually annotated by specialists, being composed of 7,000 documents annotated according to three different layers: a binary classification (offensive versus non-offensive comments), offense-level classes (highly, moderately, and slightly offensive messages), as well as nine hate speech targets (xenophobia, racism, homophobia, sexism, religious intolerance, partyism, apology to the dictatorship, antisemitism, and fatphobia). Each comment was annotated by three different annotators and achieved high inter-annotator agreement.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Abusive LanguageBinary Classification

Similar Papers 제목 키워드 기반

HateBR: A Large Expert Annotated Corpus of Brazilian Instagram Comments for Offensive Language and Hate Speech Detection

2021-03-27 · LREC 2022 6 · Francielle Alves Vargas, Isabelle Carvalho, Fabiana Rodrigues de Góes, Fabrício Benevenuto 외

Due to the severity of the social media offensive and hateful comments in Brazil, and the lack of research in Portuguese, this paper provides the first large-scale expert annotated corpus of Brazilian Instagram comments …

BIG-bench Machine LearningBinary ClassificationHate Speech Detection

Propbank-Br: a Brazilian Treebank annotated with semantic role labels

2012-05-01 · LREC 2012 5 · Magali Sanches Duran, S Alu{\'\i}sio, ra Maria

This paper reports the annotation of a Brazilian Portuguese Treebank with semantic role labels following Propbank guidelines. A different language and a different parser output impact the task and require some decisions …

Machine TranslationQuestion AnsweringSemantic Role Labeling

Self-Explaining Hate Speech Detection with Moral Rationales

2026-01-07 · Francielle Vargas, Jackson Trager, Diego Alves, Surendrabikram Thapa 외 arxiv

Hate speech detection models rely on surface-level lexical features, increasing vulnerability to spurious correlations and limiting robustness, cultural contextualization, and interpretability. We propose Supervised Mora…

Hate Speech Detection

Towards a General Abstract Meaning Representation Corpus for Brazilian Portuguese

2019-08-01 · WS 2019 8 · Marco Antonio Sobrevilla Cabezudo, Thiago Pardo

Abstract Meaning Representation (AMR) is a recent and prominent semantic representation with good acceptance and several applications in the Natural Language Processing area. For English, there is a large annotated corpu…

Abstract Meaning Representation

Essay-BR: a Brazilian Corpus of Essays

2021-05-19 · Jeziel C. Marinho, Rafael T. Anchieta, Raimundo S. Moura

Automatic Essay Scoring (AES) is defined as the computer technology that evaluates and scores the written essays, aiming to provide computational models to grade essays either automatically or with minimal human involvem…