paper-with-me

홈 › Papers

Chaos with Keywords: Exposing Large Language Models Sycophantic Hallucination to Misleading Keywords and Evaluating Defense Strategies

2024-06-06 · Aswin RRV, Nemika Tyagi, Md Nayem Uddin, Neeraj Varshney, Chitta Baral

This study explores the sycophantic tendencies of Large Language Models (LLMs), where these models tend to provide answers that match what users want to hear, even if they are not entirely correct. The motivation behind this exploration stems from the common behavior observed in individuals searching the internet for facts with partial or misleading knowledge. Similar to using web search engines, users may recall fragments of misleading keywords and submit them to an LLM, hoping for a comprehensive response. Our empirical analysis of several LLMs shows the potential danger of these models amplifying misinformation when presented with misleading keywords. Additionally, we thoroughly assess four existing hallucination mitigation strategies to reduce LLMs sycophantic behavior. Our experiments demonstrate the effectiveness of these strategies for generating factually correct statements. Furthermore, our analyses delve into knowledge-probing experiments on factual keywords and different categories of sycophancy mitigation.

📄 PDF Abstract BibTeX arXiv:2406.03827

Code (0)

등록된 구현이 없습니다.

Tasks

HallucinationKnowledge ProbingMisinformation

Similar Papers 제목 키워드 기반

Pointing to a Llama and Call it a Camel: On the Sycophancy of Multimodal Large Language Models

2025-09-19 · Renjie Pi, Kehao Miao, Li Peihang, Runtao Liu 외 arxiv

Multimodal large language models (MLLMs) have demonstrated extraordinary capabilities in conducting conversations based on image inputs. However, we observe that MLLMs exhibit a pronounced form of visual sycophantic beha…

Sycophancy Is Not One Thing: Causal Separation of Sycophantic Behaviors in LLMs

2025-09-25 · Daniel Vennemeyer, Phan Anh Duong, Tiffany Zhan, Tianyu Jiang arxiv

Large language models (LLMs) often exhibit sycophantic behaviors -- such as excessive agreement with or flattery of the user -- but it is unclear whether these behaviors arise from a single mechanism or multiple distinct…

Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model

2024-12-03 · María Victoria Carro

Sycophancy refers to the tendency of a large language model to align its outputs with the user's perceived preferences, beliefs, or opinions, in order to look favorable, regardless of whether those statements are factual…

Language ModelingLanguage ModellingLarge Language ModelMisinformation

Complacent, Not Sycophantic: Reframing Large Language Models and Designing AI Literacy for Complacent Machines

2026-05-14 · Federico Germani, Giovanni Spitale arxiv

Large language models are often described as sycophantic, in the sense that they appear to flatter users or mirror their beliefs. We argue that this label is conceptually misleading: sycophancy implies motives and strate…

When Large Language Models contradict humans? Large Language Models' Sycophantic Behaviour

2023-11-15 · Leonardo Ranaldi, Giulia Pucci

Large Language Models have been demonstrating the ability to solve complex tasks by delivering answers that are positively evaluated by humans due in part to the intensive use of human feedback that refines responses. Ho…