paper-with-me

홈 › Papers

Are Large Language Models Sensitive to the Motives Behind Communication?

2025-10-22 · Addison J. Wu, Ryan Liu, Kerem Oktar, Theodore R. Sumers, Thomas L. Griffiths arxiv

Human communication is motivated: people speak, write, and create content with a particular communicative intent in mind. As a result, information that large language models (LLMs) and AI agents process is inherently framed by humans' intentions and incentives. People are adept at navigating such nuanced information: we routinely identify benevolent or self-serving motives in order to decide what statements to trust. For LLMs to be effective in the real world, they too must critically evaluate content by factoring in the motivations of the source -- for instance, weighing the credibility of claims made in a sales pitch. In this paper, we undertake a comprehensive study of whether LLMs have this capacity for motivational vigilance. We first employ controlled experiments from cognitive science to verify that LLMs' behavior is consistent with rational models of learning from motivated testimony, and find they successfully discount information from biased sources in a human-like manner. We then extend our evaluation to sponsored online adverts, a more naturalistic reflection of LLM agents' information ecosystems. In these settings, we find that LLMs' inferences do not track the rational models' predictions nearly as closely -- partly due to additional information that distracts them from vigilance-relevant considerations. However, a simple steering intervention that boosts the salience of intentions and incentives substantially increases the correspondence between LLMs and the rational model. These results suggest that LLMs possess a basic sensitivity to the motivations of others, but generalizing to novel real-world settings will require further improvements to these models.

📄 PDF Abstract BibTeX arXiv:2510.19687

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ChatGPT Role-play Dataset: Analysis of User Motives and Model Naturalness

2024-03-26 · Yufei Tao, Ameeta Agrawal, Judit Dombi, Tetyana Sydorenko 외

Recent advances in interactive large language models like ChatGPT have revolutionized various domains; however, their behavior in natural and role-play conversation settings remains underexplored. In our study, we addres…

Diversity

Observing Micromotives and Macrobehavior of Large Language Models

2024-12-10 · Yuyang Cheng, Xingwei Qu, Tomas Goldsack, Chenghua Lin 외

Thomas C. Schelling, awarded the 2005 Nobel Memorial Prize in Economic Sciences, pointed out that ``individuals decisions (micromotives), while often personal and localized, can lead to societal outcomes (macrobehavior) …

Toward Comprehensive Understanding of a Sentiment Based on Human Motives

2019-07-01 · ACL 2019 7 · Naoki Otani, Eduard Hovy

In sentiment detection, the natural language processing community has focused on determining holders, facets, and valences, but has paid little attention to the reasons for sentiment decisions. Our work considers human m…

Transfer Learning

Learning to request guidance in emergent language

2019-11-01 · WS 2019 11 · Benjamin Kolb, Leon Lang, Henning Bartsch, Arwin Gansekoele 외

Previous research into agent communication has shown that a pre-trained guide can speed up the learning process of an imitation learning agent. The guide achieves this by providing the agent with discrete messages in an …

Imitation Learning

Learning to Request Guidance in Emergent Communication

2019-12-11 · Benjamin Kolb, Leon Lang, Henning Bartsch, Arwin Gansekoele 외

Previous research into agent communication has shown that a pre-trained guide can speed up the learning process of an imitation learning agent. The guide achieves this by providing the agent with discrete messages in an …

Imitation Learning