paper-with-me

홈 › Papers

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

2026-02-06 · Gaia Molinaro, Dave August, Danielle Perszyk, Anne G. E. Collins arxiv

Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in choosing which goals to pursue, rather than completing predefined tasks. However, the assumption that LLMs accurately reflect human preferences for goal setting remains largely untested. We assess the validity of LLMs as proxies for human goal selection in a controlled, self-directed learning task borrowed from cognitive science. Across five models (GPT-5, Gemini 2.5 Pro, Claude Sonnet 4.5, Qwen3 32B, and Centaur), we find substantial divergence from human behavior. While people gradually explore and learn to achieve goals with diversity across individuals, most models exploit a single identified solution or show surprisingly low performance, with distinct patterns across models and little variability across instances of the same model. Chain-of-thought reasoning and persona steering provide limited improvements, and our conclusions hold across experimental settings. While they await confirmation in applied settings, these findings highlight the uniqueness of human goal selection and caution against its replacement with current models.

📄 PDF Abstract BibTeX arXiv:2603.03295

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

How to Count AIs: Individuation and Liability for AI Agents

2026-02-24 · Yonathan Arbel, Peter Salib, Simon Goldstein arxiv

Very soon, millions of AI agents will proliferate across the economy, autonomously taking billions of actions. Inevitably, things will go wrong. Humans will be defrauded, injured, even killed. Law will somehow have to go…

Modeling Human Inference of Others' Intentions in Complex Situations with Plan Predictability Bias

2018-05-16 · Ryo Nakahashi, Seiji Yamada

A recent approach based on Bayesian inverse planning for the "theory of mind" has shown good performance in modeling human cognition. However, perfect inverse planning differs from human cognition during one kind of comp…

Incremental Neural Lexical Coherence Modeling

2020-12-01 · COLING 2020 8 · Sungho Jeon, Michael Strube

Pretrained language models, neural models pretrained on massive amounts of data, have established the state of the art in a range of NLP tasks. They are based on a modern machine-learning technique, the Transformer which…

Language ModelingLanguage Modelling

Augmenting Autotelic Agents with Large Language Models

2023-05-21 · Cédric Colas, Laetitia Teodorescu, Pierre-Yves Oudeyer, Xingdi Yuan 외

Humans learn to master open-ended repertoires of skills by imagining and practicing their own goals. This autotelic learning process, literally the pursuit of self-generated (auto) goals (telos), becomes more and more op…

Common Sense ReasoningLanguage ModelingLanguage Modelling

DetectGPT-SC: Improving Detection of Text Generated by Large Language Models through Self-Consistency with Masked Predictions

2023-10-23 · Rongsheng Wang, Qi Li, Sihong Xie

General large language models (LLMs) such as ChatGPT have shown remarkable success, but it has also raised concerns among people about the misuse of AI-generated texts. Therefore, an important question is how to detect w…

Logical ReasoningText Generation