paper-with-me

Papers

MuTox: Universal MUltilingual Audio-based TOXicity Dataset and Zero-shot Detector

2024-01-10 · Marta R. Costa-jussà, Mariano Coria Meglioli, Pierre Andrews, David Dale, Prangthip Hansanti, Elahe Kalbassi, Alex Mourachko, Christophe Ropers, Carleigh Wood

Research in toxicity detection in natural language processing for the speech modality (audio-based) is quite limited, particularly for languages other than English. To address these limitations and lay the groundwork for truly multilingual audio-based toxicity detection, we introduce MuTox, the first highly multilingual audio-based dataset with toxicity labels. The dataset comprises 20,000 audio utterances for English and Spanish, and 4,000 for the other 19 languages. To demonstrate the quality of this dataset, we trained the MuTox audio-based toxicity classifier, which enables zero-shot toxicity detection across a wide range of languages. This classifier outperforms existing text-based trainable classifiers by more than 1% AUC, while expanding the language coverage more than tenfold. When compared to a wordlist-based classifier that covers a similar number of languages, MuTox improves precision and recall by approximately 2.5 times. This significant improvement underscores the potential of MuTox in advancing the field of audio-based toxicity detection.

📄 PDF Abstract BibTeX arXiv:2401.05060

Code (1)

facebookresearch/seamless_communication 공식 구현 pytorch

Similar Papers 제목 키워드 기반

On the Role of Speech Data in Reducing Toxicity Detection Bias

2024-11-12 · Samuel J. Bell, Mariano Coria Meglioli, Megan Richards, Eduardo Sánchez 외

Text toxicity detection systems exhibit significant biases, producing disproportionate rates of false positives on samples mentioning demographic groups. But what about toxicity detection in speech? To investigate the ex…

Enhancing Multilingual Voice Toxicity Detection with Speech-Text Alignment

2024-06-14 · Joseph Liu, Mahesh Kumar Nandwana, Janne Pylkkönen, Hannes Heikinheimo 외

Toxicity classification for voice heavily relies on the semantic content of speech. We propose a novel framework that utilizes cross-modal learning to integrate the semantic embedding of text into a multilabel speech tox…

Classification

MultiLinguahah : A New Unsupervised Multilingual Acoustic Laughter Segmentation Method

2026-05-07 · Sofia Callejas, Nahuel Gomez, Catherine Pelachaud, Brian Ravenet 외 arxiv

Laughter is a social non-vocalization that is universal across cultures and languages, and is crucial for human communication, including social bonding and communication signaling. However, detecting laughter in audio is…

Anomaly Detection

PolygloToxicityPrompts: Multilingual Evaluation of Neural Toxic Degeneration in Large Language Models

2024-05-15 · Devansh Jain, Priyanshu Kumar, Samuel Gehman, Xuhui Zhou 외

Recent advances in large language models (LLMs) have led to their extensive global deployment, and ensuring their safety calls for comprehensive and multilingual toxicity evaluations. However, existing toxicity benchmark…

Benchmarking

Maya: An Instruction Finetuned Multilingual Multimodal Model

2024-12-10 · Nahid Alam, Karthik Reddy Kanjula, Surya Guthikonda, Timothy Chung 외

The rapid development of large Vision-Language Models (VLMs) has led to impressive results on academic benchmarks, primarily in widely spoken languages. However, significant gaps remain in the ability of current VLMs to …

model