paper-with-me

홈 › Papers

Jibes & Delights: A Dataset of Targeted Insults and Compliments to Tackle Online Abuse

2021-08-01 · ACL (WOAH) 2021 8 · Ravsimar Sodhi, Kartikey Pant, Radhika Mamidi

Online abuse and offensive language on social media have become widespread problems in today’s digital age. In this paper, we contribute a Reddit-based dataset, consisting of 68,159 insults and 51,102 compliments targeted at individuals instead of targeting a particular community or race. Secondly, we benchmark multiple existing state-of-the-art models for both classification and unsupervised style transfer on the dataset. Finally, we analyse the experimental results and conclude that the transfer task is challenging, requiring the models to understand the high degree of creativity exhibited in the data.

📄 PDF Abstract BibTeX

Code (1)

ravsimar-sodhi/jibes-and-delights 공식 구현

Tasks

Style Transfer

Similar Papers 제목 키워드 기반

Turkish Delights: a Dataset on Turkish Euphemisms

2024-07-17 · Hasan Can Biyik, Patrick Lee, Anna Feldman

Euphemisms are a form of figurative language relatively understudied in natural language processing. This research extends the current computational work on potentially euphemistic terms (PETs) to Turkish. We introduce t…

Binary Classification

Neural operator learning of heterogeneous mechanobiological insults contributing to aortic aneurysms

2022-05-08 · Somdatta Goswami, David S. Li, Bruno V. Rego, Marcos Latorre 외

Thoracic aortic aneurysm (TAA) is a localized dilatation of the aorta resulting from compromised wall composition, structure, and function, which can lead to life-threatening dissection or rupture. Several genetic mutati…

Operator learning

Neural Word Decomposition Models for Abusive Language Detection

2019-10-02 · WS 2019 8 · Sravan Babu Bodapati, Spandana Gella, Kasturi Bhattacharjee, Yaser Al-Onaizan

User generated text on social media often suffers from a lot of undesired characteristics including hatespeech, abusive language, insults etc. that are targeted to attack or abuse a specific group of people. Often such t…

Abusive Language

Overview of OSACT4 Arabic Offensive Language Detection Shared Task

2020-05-01 · LREC 2020 5 · Hamdy Mubarak, Kareem Darwish, Walid Magdy, Tamer Elsayed 외

This paper provides an overview of the offensive language detection shared task at the 4th workshop on Open-Source Arabic Corpora and Processing Tools (OSACT4). There were two subtasks, namely: Subtask A, involving the d…

Diagnosing and Debiasing Corpus-Based Political Bias and Insults in GPT2

2023-11-17 · Ambri Ma, Arnav Kumar, Brett Zeligson

The training of large language models (LLMs) on extensive, unfiltered corpora sourced from the internet is a common and advantageous practice. Consequently, LLMs have learned and inadvertently reproduced various types of…