paper-with-me

홈 › Papers

Teaching Values to Machines: Simulating Human-Like Behavior in LLMs

2026-05-28 · Asaf Yehudai, Naama Rozen, Ariel Gera arxiv

Large Language Models (LLMs) demonstrate a remarkable capacity to adopt different personas and roles; however, it remains unclear whether they can manifest behavior that adheres to a coherent, human-like value structure. In this work, we draw on established psychological value theory to induce human-like values in LLMs and assess their alignment with patterns observed in human studies. Using validated psychological questionnaires, we conduct large-scale experiments -- over 5 million questions -- to evaluate value structures and value-behavior relationships in leading LLMs and compare them to humans. Our findings reveal strong agreement between value-prompted LLMs and humans across both dimensions. Moreover, incorporating human value distributions enhances population-level simulations with value-induced LLMs. These findings highlight the potential of value-induced LLMs as effective, psychologically grounded tools for simulating human behavior.

📄 PDF Abstract BibTeX arXiv:2605.30036

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Computational Narrative Intelligence: A Human-Centered Goal for Artificial Intelligence

2016-02-21 · Mark O. Riedl

Narrative intelligence is the ability to craft, tell, understand, and respond affectively to stories. We argue that instilling artificial intelligences with computational narrative intelligence affords a number of applic…

BIG-bench Machine Learning

One-shot Machine Teaching: Cost Very Few Examples to Converge Faster

2022-12-13 · Chen Zhang, Xiaofeng Cao, Yi Chang, Ivor W Tsang

Artificial intelligence is to teach machines to take actions like humans. To achieve intelligent teaching, the machine learning community becomes to think about a promising topic named machine teaching where the teacher …

We Urgently Need Intrinsically Kind Machines

2024-10-21 · Joshua T. S. Hewson

Artificial Intelligence systems are rapidly evolving, integrating extrinsic and intrinsic motivations. While these frameworks offer benefits, they risk misalignment at the algorithmic level while appearing superficially …

Teaching is a Process: The TOSS Framework for Modeling Human Teaching Decisions in Human-Interactive Robot Learning

2026-08-21 · Bernhard Hilpert, Kim Baraka, Joost Broekens arxiv

Successful Human-Robot Teaching assumes alignment between robot processing needs and human teaching intent. To better understand this alignment, this work seeks to uncover the underlying logic that humans intuitively app…

Reinforcement Learning

Toward a general, scaleable framework for Bayesian teaching with applications to topic models

2016-05-25 · Baxter S. Eaves Jr, Patrick Shafto

Machines, not humans, are the world's dominant knowledge accumulators but humans remain the dominant decision makers. Interpreting and disseminating the knowledge accumulated by machines requires expertise, time, and is …

Topic Models