paper-with-me

홈 › Papers

Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models

2026-04-09 · Yousra Fettach, Guillaume Bied, Hannu Toivonen, Tijl De Bie arxiv

Humor is one of the most culturally embedded and socially significant dimensions of human communication, yet it remains largely unexplored as a dimension of Large Language Model (LLM) alignment. In this study, five frontier language models play the same Cards Against Humanity games (CAH) as human players. The models select the funniest response from a slate of ten candidate cards across 9,894 rounds. While all models exceed the random baseline, alignment with human preference remains modest. More striking is that models agree with each other substantially more often than they agree with humans. We show that this preference is partly explained by systematic position biases and content preferences, raising the question whether LLM humor judgment reflects genuine preference or structural artifacts of inference and alignment.

📄 PDF Abstract BibTeX arXiv:2604.08757

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Cards Against AI: Predicting Humor in a Fill-in-the-blank Party Game

2022-10-24 · Dan Ofer, Dafna Shahaf

Humor is an inherently social phenomenon, with humorous utterances shaped by what is socially and culturally accepted. Understanding humor is an important NLP challenge, with many applications to human-computer interacti…

Feature Importance

HumorReject: Decoupling LLM Safety from Refusal Prefix via A Little Humor

2025-01-23 · Zihui Wu, Haichang Gao, Jiacheng Luo, Zhaoxiang Liu

Large Language Models (LLMs) commonly rely on explicit refusal prefixes for safety, making them vulnerable to prefix injection attacks. We introduce HumorReject, a novel data-driven approach that reimagines LLM safety by…

Chumor 2.0: Towards Benchmarking Chinese Humor Understanding

2024-12-23 · Ruiqi He, Yushu He, Longju Bai, Jiarui Liu 외

Existing humor datasets and evaluations predominantly focus on English, leaving limited resources for culturally nuanced humor in non-English languages like Chinese. To address this gap, we construct Chumor, the first Ch…

Benchmarking

Chumor 1.0: A Truly Funny and Challenging Chinese Humor Understanding Dataset from Ruo Zhi Ba

2024-06-18 · Ruiqi He, Yushu He, Longju Bai, Jiarui Liu 외

Existing humor datasets and evaluations predominantly focus on English, lacking resources for culturally nuanced humor in non-English languages like Chinese. To address this gap, we construct Chumor, a dataset sourced fr…

Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges

2026-07-21 · Tuo Liang, Zhe Hu, Disheng Liu, Jing Li 외 hf

Multimodal humor in memes, cartoons, and comics remains difficult for AI systems because intended meaning depends on non-literal mechanisms, shared cultural knowledge, and communicative intent rather than literal scene d…