paper-with-me

Papers

Bridging Modalities: Enhancing Cross-Modality Hate Speech Detection with Few-Shot In-Context Learning

2024-10-08 · Ming Shan Hee, Aditi Kumaresan, Roy Ka-Wei Lee

The widespread presence of hate speech on the internet, including formats such as text-based tweets and vision-language memes, poses a significant challenge to digital platform safety. Recent research has developed detection models tailored to specific modalities; however, there is a notable gap in transferring detection capabilities across different formats. This study conducts extensive experiments using few-shot in-context learning with large language models to explore the transferability of hate speech detection between modalities. Our findings demonstrate that text-based hate speech examples can significantly enhance the classification accuracy of vision-language hate speech. Moreover, text-based demonstrations outperform vision-language demonstrations in few-shot learning settings. These results highlight the effectiveness of cross-modality knowledge transfer and offer valuable insights for improving hate speech detection systems.

📄 PDF Abstract BibTeX arXiv:2410.05600

Code (0)

등록된 구현이 없습니다.

Tasks

Few-Shot LearningHate Speech DetectionIn-Context LearningTransfer Learning

Similar Papers 제목 키워드 기반

Towards a Robust Framework for Multimodal Hate Detection: A Study on Video vs. Image-based Content

2025-02-11 · Girish A. Koushik, Diptesh Kanojia, Helen Treharne

Social media platforms enable the propagation of hateful content across different modalities such as textual, auditory, and visual, necessitating effective detection methods. While recent approaches have shown promise in…

Hate Speech DetectionVideo Classification

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion

2025-05-17 · Yinghui Zhang, Tailin Chen, Yuchen Zhang, Zeyu Fu

The rapid rise of video content on platforms such as TikTok and YouTube has transformed information dissemination, but it has also facilitated the spread of harmful content, particularly hate videos. Despite significant …

Zero-resource Machine Translation by Multimodal Encoder-decoder Network with Multimedia Pivot

2016-11-14 · Hideki Nakayama, Noriki Nishida

We propose an approach to build a neural machine translation system with no supervised resources (i.e., no parallel corpora) using multimodal embedded representation over texts and images. Based on the assumption that te…

DecoderMachine TranslationTranslation

EventDance++: Language-guided Unsupervised Source-free Cross-modal Adaptation for Event-based Object Recognition

2024-09-19 · Xu Zheng, Lin Wang

In this paper, we address the challenging problem of cross-modal (image-to-events) adaptation for event-based recognition without accessing any labeled source image data. This task is arduous due to the substantial modal…

Object Recognition

Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach

2024-08-14 · Muhammad Saad Saeed, Shah Nawaz, Muhammad Zaigham Zaheer, Muhammad Haris Khan 외

Multimodal networks have demonstrated remarkable performance improvements over their unimodal counterparts. Existing multimodal networks are designed in a multi-branch fashion that, due to the reliance on fusion strategi…