paper-with-me

Papers

D-HUMOR: Dark Humor Understanding via Multimodal Open-ended Reasoning -- A Benchmark Dataset and Method

2025-09-08 · Sai Kartheek Reddy Kasu, Mohammad Zia Ur Rehman, Shahid Shafi Dar, Rishi Bharat Junghare, Dhanvin Sanjay Namboodiri, Nagendra Kumar arxiv

Dark humor in online memes poses unique challenges due to its reliance on implicit, sensitive, and culturally contextual cues. To address the lack of resources and methods for detecting dark humor in multimodal content, we introduce a novel dataset of 4,379 Reddit memes annotated for dark humor, target category (gender, mental health, violence, race, disability, and other), and a three-level intensity rating (mild, moderate, severe). Building on this resource, we propose a reasoning-augmented framework that first generates structured explanations for each meme using a Large Vision-Language Model (VLM). Through a Role-Reversal Self-Loop, VLM adopts the author's perspective to iteratively refine its explanations, ensuring completeness and alignment. We then extract textual features from both the OCR transcript and the self-refined reasoning via a text encoder, while visual features are obtained using a vision transformer. A Tri-stream Cross-Reasoning Network (TCRNet) fuses these three streams, text, image, and reasoning, via pairwise attention mechanisms, producing a unified representation for classification. Experimental results demonstrate that our approach outperforms strong baselines across three tasks: dark humor detection, target identification, and intensity prediction. The dataset, annotations, and code are released to facilitate further research in multimodal humor understanding and content moderation. Code and Dataset are available at: https://github.com/Sai-Kartheek-Reddy/D-Humor-Dark-Humor-Understanding-via-Multimodal-Open-ended-Reasoning

📄 PDF Abstract BibTeX arXiv:2509.06771

Code (0)

등록된 구현이 없습니다.

Tasks

Dark Humor Detection

Similar Papers 제목 키워드 기반

Harm or Humor: A Multimodal, Multilingual Benchmark for Overt and Covert Harmful Humor

2026-03-18 · Ahmed Sharshar, Hosam Elgendy, Saad El Dine Ahmed, Yasser Rohaim 외 arxiv

Dark humor often relies on subtle cultural nuances and implicit cues that require contextual reasoning to interpret, posing safety challenges that current static benchmarks fail to capture. To address this, we introduce …

UR-FUNNY: A Multimodal Language Dataset for Understanding Humor

2019-04-14 · IJCNLP 2019 11 · Md. Kamrul Hasan, Wasifur Rahman, Amir Zadeh, Jianyuan Zhong 외

Humor is a unique and creative communicative behavior displayed during social interactions. It is produced in a multimodal manner, through the usage of words (text), gestures (vision) and prosodic cues (acoustic). Unders…

Humor Detection

When Jokes Cross the Line: Analyzing Regular Humor and Dark Humor in YouTube Shorts

2026-04-30 · Sydney Johns, Sanjeev Parthasarathy, Shantnu Bhalla, Vaibhav Garg arxiv

Video platforms such as YouTube have reshaped how users engage with entertainment and information, emphasizing brief, highly engaging content such as Shorts. Within this ecosystem, certain content occupies a gray area wh…

v-HUB: A Benchmark for Video Humor Understanding from Vision and Sound

2025-09-30 · Zhengpeng Shi, Yanpeng Zhao, Jianqun Zhou, Yuxuan Wang 외 arxiv

AI models capable of comprehending humor hold real-world promise -- for example, enhancing engagement in human-machine interactions. To gauge and diagnose the capacity of multimodal large language models (MLLMs) for humo…

Is AI fun? HumorDB: a curated dataset and benchmark to investigate graphical humor

2024-06-19 · Veedant Jain, Felipe dos Santos Alves Feitosa, Gabriel Kreiman

Despite significant advancements in computer vision, understanding complex scenes, particularly those involving humor, remains a substantial challenge. This paper introduces HumorDB, a novel image-only dataset specifical…

Binary Classification