paper-with-me

Papers

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

2025-03-29 · Tuo Liang, Zhe Hu, Jing Li, Hao Zhang, Yiren Lu, Yunlai Zhou, Yiran Qiao, Disheng Liu, Jeirui Peng, Jing Ma, Yu Yin

Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant challenge for large vision-language models (VLMs). This limitation hinders AI's ability to engage in human-like reasoning and cultural expression. In this paper, we investigate this challenge through an in-depth analysis of comics that juxtapose panels to create humor through contradictions. We introduce the YesBut (V2), a novel benchmark with 1,262 comic images from diverse multilingual and multicultural contexts, featuring comprehensive annotations that capture various aspects of narrative understanding. Using this benchmark, we systematically evaluate a wide range of VLMs through four complementary tasks spanning from surface content comprehension to deep narrative reasoning, with particular emphasis on comparative reasoning between contradictory elements. Our extensive experiments reveal that even the most advanced models significantly underperform compared to humans, with common failures in visual perception, key element identification, comparative analysis and hallucinations. We further investigate text-based training strategies and social knowledge augmentation methods to enhance model performance. Our findings not only highlight critical weaknesses in VLMs' understanding of cultural and creative expressions but also provide pathways toward developing context-aware models capable of deeper narrative understanding though comparative reasoning.

📄 PDF Abstract BibTeX arXiv:2503.23137

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions

2024-05-29 · Zhe Hu, Tuo Liang, Jing Li, Yiren Lu 외

Recent advancements in large multimodal language models have demonstrated remarkable proficiency across a wide range of tasks. Yet, these models still struggle with understanding the nuances of human humor through juxtap…

v-HUB: A Benchmark for Video Humor Understanding from Vision and Sound

2025-09-30 · Zhengpeng Shi, Yanpeng Zhao, Jianqun Zhou, Yuxuan Wang 외 arxiv

AI models capable of comprehending humor hold real-world promise -- for example, enhancing engagement in human-machine interactions. To gauge and diagnose the capacity of multimodal large language models (MLLMs) for humo…

Exploring Chinese Humor Generation: A Study on Two-Part Allegorical Sayings

2024-03-16 · Rongwu Xu

Humor, a culturally nuanced aspect of human language, poses challenges for computational understanding and generation, especially in Chinese humor, which remains relatively unexplored in the NLP community. This paper inv…

Contrastive LearningLanguage ModelingLanguage Modelling

DuanzAI: Slang-Enhanced LLM with Prompt for Humor Understanding

2024-05-23 · Yesian Rohn

Language's complexity is evident in the rich tapestry of slang expressions, often laden with humor and cultural nuances. This linguistic phenomenon has become increasingly prevalent, especially in digital communication. …

Chatbot

Dutch Humor Detection by Generating Negative Examples

2020-10-26 · Thomas Winters, Pieter Delobelle

Detecting if a text is humorous is a hard task to do computationally, as it usually requires linguistic and common sense insights. In machine learning, humor detection is usually modeled as a binary classification task, …

Binary ClassificationCommon Sense ReasoningHumor DetectionLanguage Modelling+1