paper-with-me

Papers

Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation

2026-01-10 · Jen-tse Huang, Chang Chen, Shiyang Lai, Wenxuan Wang, Michelle R. Kaufman, Mark Dredze arxiv

Short-video platforms have become major channels for misinformation, where deceptive claims frequently leverage visual experiments and social cues. While Multimodal Large Language Models (MLLMs) have demonstrated impressive reasoning capabilities, their robustness against misinformation entangled with cognitive biases remains under-explored. In this paper, we introduce a comprehensive evaluation framework using a high-quality, manually annotated dataset of 200 short videos spanning four health domains. This dataset provides fine-grained annotations for three deceptive patterns-experimental errors, logical fallacies, and fabricated claims-each verified by evidence such as national standards and academic literature. We evaluate eight frontier MLLMs across five modality settings. Experimental results demonstrate that Gemini-2.5-Pro achieves the highest performance in the multimodal setting with a belief score of 71.5/100, while o3 performs the worst at 35.2. Furthermore, we investigate social cues that induce false beliefs in videos and find that models are susceptible to biases like authoritative channel IDs.

📄 PDF Abstract BibTeX arXiv:2601.06600

Code (0)

등록된 구현이 없습니다.

Tasks

Logical Fallacies

Similar Papers 제목 키워드 기반

Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding

2025-08-28 · Vanessa Figueiredo arxiv

We study how prompt-level inductive biases influence the cognitive behavior of large language models (LLMs) in instructional dialogue. We introduce a symbolic scaffolding method paired with a short-term memory schema des…

A Comprehensive Evaluation of Cognitive Biases in LLMs

2024-10-20 · Simon Malberg, Roman Poletukhin, Carolin M. Schuster, Georg Groh

We present a large-scale evaluation of 30 cognitive biases in 20 state-of-the-art large language models (LLMs) under various decision-making scenarios. Our contributions include a novel general-purpose test framework for…

Decision Making

Probing Association Biases in LLM Moderation Over-Sensitivity

2025-05-29 · Yuxin Wang, Botao Yu, Ivory Yang, Saeed Hassanpour 외

Large Language Models are widely used for content moderation but often misclassify benign comments as toxic, leading to over-sensitivity. While previous research attributes this issue primarily to the presence of offensi…

Sensitivity

Psychology of Artificial Intelligence: Epistemological Markers of the Cognitive Analysis of Neural Networks

2024-07-04 · Michael Pichat

What is the "nature" of the cognitive processes and contents of an artificial neural network? In other words, how does an artificial intelligence fundamentally "think," and in what form does its knowledge reside? The psy…

Attribute

Cognitive bias in large language models: Cautious optimism meets anti-Panglossian meliorism

2023-11-18 · David Thorstad

Traditional discussions of bias in large language models focus on a conception of bias closely tied to unfairness, especially as affecting marginalized groups. Recent work raises the novel possibility of assessing the ou…