paper-with-me

Papers

Do Large Language Models Exhibit Cognitive Dissonance? Studying the Difference Between Revealed Beliefs and Stated Answers

2024-06-21 · Manuel Mondal, Ljiljana Dolamic, Gérôme Bovet, Philippe Cudré-Mauroux, Julien Audiffren

Prompting and Multiple Choices Questions (MCQ) have become the preferred approach to assess the capabilities of Large Language Models (LLMs), due to their ease of manipulation and evaluation. Such experimental appraisals have pointed toward the LLMs' apparent ability to perform causal reasoning or to grasp uncertainty. In this paper, we investigate whether these abilities are measurable outside of tailored prompting and MCQ by reformulating these issues as direct text completion - the foundation of LLMs. To achieve this goal, we define scenarios with multiple possible outcomes and we compare the prediction made by the LLM through prompting (their Stated Answer) to the probability distributions they compute over these outcomes during next token prediction (their Revealed Belief). Our findings suggest that the Revealed Belief of LLMs significantly differs from their Stated Answer and hint at multiple biases and misrepresentations that their beliefs may yield in many scenarios and outcomes. As text completion is at the core of LLMs, these results suggest that common evaluation methods may only provide a partial picture and that more research is needed to assess the extent and nature of their capabilities.

📄 PDF Abstract BibTeX arXiv:2406.14986

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

HINT An unsupervised approach for identifying Hierarchical Information Threads by analysing the network of related articles in a collection. In particular, HINT leverages article…

Similar Papers 제목 키워드 기반

Confident but Conflicted: Internal Uncertainty and Cognitive Dissonance Resolution in LLMs

2026-06-21 · Weihong Qi, Kristina Lerman arxiv

Large language models (LLMs) frequently encounter inputs that disagree with their prior outputs, through user pushback, retrieved documents, or web search results. While the way they resolve such conflicts -- a process w…

Characterizing Social Imaginaries and Self-Disclosures of Dissonance in Online Conspiracy Discussion Communities

2021-07-21 · Shruti Phadke, Mattia Samory, Tanushree Mitra

Online discussion platforms offer a forum to strengthen and propagate belief in misinformed conspiracy theories. Yet, they also offer avenues for conspiracy theorists to express their doubts and experiences of cognitive …

2k

The Cognitive Circuit Breaker: A Systems Engineering Framework for Intrinsic AI Reliability

2026-04-15 · Jonathan Pan arxiv

As Large Language Models (LLMs) are increasingly deployed in mission-critical software systems, detecting hallucinations and ``faked truthfulness'' has become a paramount engineering challenge. Current reliability archit…

Dissecting Dissonance: Benchmarking Large Multimodal Models Against Self-Contradictory Instructions

2024-08-02 · Jin Gao, Lei Gan, Yuankai Li, Yixin Ye 외

Large multimodal models (LMMs) excel in adhering to human instructions. However, self-contradictory instructions may arise due to the increasing trend of multimodal interaction and context length, which is challenging fo…

Benchmarkingmultimodal interaction

Topic-independent Detection of Dissonance in Short Stance Text

2021-06-16 · ACL ARR Jun 2021 6 · Anonymous

We address dissonance detection, the task of detecting conflicting stance between two input statements. Computational models for stance detection have typically been trained for a given target topic (e.g. gun control). I…

Stance Detection