paper-with-me

홈 › Papers

Grounding Toxicity in Real-World Events across Languages

2024-05-22 · Wondimagegnhue Tsegaye Tufa, Ilia Markov, Piek Vossen

Social media conversations frequently suffer from toxicity, creating significant issues for users, moderators, and entire communities. Events in the real world, like elections or conflicts, can initiate and escalate toxic behavior online. Our study investigates how real-world events influence the origin and spread of toxicity in online discussions across various languages and regions. We gathered Reddit data comprising 4.5 million comments from 31 thousand posts in six different languages (Dutch, English, German, Arabic, Turkish and Spanish). We target fifteen major social and political world events that occurred between 2020 and 2023. We observe significant variations in toxicity, negative sentiment, and emotion expressions across different events and language communities, showing that toxicity is a complex phenomenon in which many different factors interact and still need to be investigated. We will release the data for further research along with our code.

📄 PDF Abstract BibTeX arXiv:2405.13754

Code (1)

cltl/grounding-toxicity 공식 구현

Similar Papers 제목 키워드 기반

Enhancing Multilingual Voice Toxicity Detection with Speech-Text Alignment

2024-06-14 · Joseph Liu, Mahesh Kumar Nandwana, Janne Pylkkönen, Hannes Heikinheimo 외

Toxicity classification for voice heavily relies on the semantic content of speech. We propose a novel framework that utilizes cross-modal learning to integrate the semantic embedding of text into a multilabel speech tox…

Classification

Learning for Dose Allocation in Adaptive Clinical Trials with Safety Constraints

2020-06-09 · ICML 2020 1 · Cong Shen, Zhiyang Wang, Sofia S. Villar, Mihaela van der Schaar

Phase I dose-finding trials are increasingly challenging as the relationship between efficacy and toxicity of new compounds (or combination of them) becomes more complex. Despite this, most commonly used methods in pract…

TimeTox: An LLM-Based Pipeline for Automated Extraction of Time Toxicity from Clinical Trial Protocols

2026-03-22 · Saketh Vinjamuri, Marielle Fis Loperena, Marie C. Spezia, Ramez Kouzy arxiv

Time toxicity, the cumulative healthcare contact days from clinical trial participation, is an important but labor-intensive metric to extract from protocol documents. We developed TimeTox, an LLM-based pipeline for auto…

Beyond Grounding: Extracting Fine-Grained Event Hierarchies Across Modalities

2022-06-14 · Hammad A. Ayyubi, Christopher Thomas, Lovish Chum, Rahul Lokesh 외

Events describe happenings in our world that are of importance. Naturally, understanding events mentioned in multimedia content and how they are related forms an important way of comprehending our world. Existing literat…

Visual Contextual Attack: Jailbreaking MLLMs with Image-Driven Context Injection

2025-07-03 · Ziqi Miao, Yi Ding, Lijun Li, Jing Shao

With the emergence of strong visual-language capabilities, multimodal large language models (MLLMs) have demonstrated tremendous potential for real-world applications. However, the security vulnerabilities exhibited by t…