paper-with-me

홈 › Papers

How Do LLMs Persuade? Linear Probes Can Uncover Persuasion Dynamics in Multi-Turn Conversations

2025-08-07 · Brandon Jaipersaud, David Krueger, Ekdeep Singh Lubana arxiv

Large Language Models (LLMs) have started to demonstrate the ability to persuade humans, yet our understanding of how this dynamic transpires is limited. Recent work has used linear probes, lightweight tools for analyzing model representations, to study various LLM skills such as the ability to model user sentiment and political perspective. Motivated by this, we apply probes to study persuasion dynamics in natural, multi-turn conversations. We leverage insights from cognitive science to train probes on distinct aspects of persuasion: persuasion success, persuadee personality, and persuasion strategy. Despite their simplicity, we show that they capture various aspects of persuasion at both the sample and dataset levels. For instance, probes can identify the point in a conversation where the persuadee was persuaded or where persuasive success generally occurs across the entire dataset. We also show that in addition to being faster than expensive prompting-based approaches, probes can do just as well and even outperform prompting in some settings, such as when uncovering persuasion strategy. This suggests probes as a plausible avenue for studying other complex behaviours such as deception and manipulation, especially in multi-turn settings and large-scale dataset analysis where prompting-based methods would be computationally inefficient.

📄 PDF Abstract BibTeX arXiv:2508.05625

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

How LLMs Are Persuaded: A Few Attention Heads, Rerouted

2026-05-10 · Xiangkun Sun, Lingkai Kong, Aoqi Zhang, Liang Zeng 외 arxiv

Language models can be persuaded to abandon factual knowledge. This vulnerability is central to AI safety, but its internal mechanism remains poorly understood. We uncover a compact causal mechanism for persuasion-induce…

Emergent Persuasion: Will LLMs Persuade Without Being Prompted?

2025-12-20 · Vincent Chang, Thee Ho, Sunishchal Dev, Kevin Zhu 외 arxiv

With the wide-scale adoption of conversational AI systems, AI are now able to exert unprecedented influence on human opinion and beliefs. Recent work has shown that many Large Language Models (LLMs) comply with requests …

Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models

2025-03-03 · Nimet Beyza Bozdag, Shuhaib Mehri, Gokhan Tur, Dilek Hakkani-Tür

Large Language Models (LLMs) demonstrate persuasive capabilities that rival human-level persuasion. While these capabilities can be used for social good, they also present risks of potential misuse. Moreover, LLMs' susce…

Misinformation

Large Language Models Are More Persuasive Than Incentivized Human Persuaders

2025-05-14 · Philipp Schoenegger, Francesco Salvi, Jiacheng Liu, Xiaoli Nan 외

We directly compare the persuasion capabilities of a frontier large language model (LLM; Claude Sonnet 3.5) against incentivized human persuaders in an interactive, real-time conversational quiz setting. In this preregis…

Language ModelingLanguage ModellingLarge Language Model

When AI Gets Persuaded, Humans Follow: Inducing the Conformity Effect in Persuasive Dialogue

2025-10-05 · Rikuo Sasaki, Michimasa Inaba arxiv

Recent advancements in AI have highlighted its application in captology, the field of using computers as persuasive technologies. We hypothesized that the "conformity effect," where individuals align with others' actions…