paper-with-me

홈 › Papers

Pragmatic-Pedagogic Value Alignment

2017-07-20 · Jaime F. Fisac, Monica A. Gates, Jessica B. Hamrick, Chang Liu, Dylan Hadfield-Menell, Malayandi Palaniappan, Dhruv Malik, S. Shankar Sastry, Thomas L. Griffiths, Anca D. Dragan

As intelligent systems gain autonomy and capability, it becomes vital to ensure that their objectives match those of their human users; this is known as the value-alignment problem. In robotics, value alignment is key to the design of collaborative robots that can integrate into human workflows, successfully inferring and adapting to their users' objectives as they go. We argue that a meaningful solution to value alignment must combine multi-agent decision theory with rich mathematical models of human cognition, enabling robots to tap into people's natural collaborative capabilities. We present a solution to the cooperative inverse reinforcement learning (CIRL) dynamic game based on well-established cognitive models of decision making and theory of mind. The solution captures a key reciprocity relation: the human will not plan her actions in isolation, but rather reason pedagogically about how the robot might learn from them; the robot, in turn, can anticipate this and interpret the human's actions pragmatically. To our knowledge, this work constitutes the first formal analysis of value alignment grounded in empirically validated cognitive models.

📄 PDF Abstract BibTeX arXiv:1707.06354

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingReinforcement Learning

Similar Papers 제목 키워드 기반

Pragmatically Learning from Pedagogical Demonstrations in Multi-Goal Environments

2022-06-09 · Hugo Caselles-Dupré, Olivier Sigaud, Mohamed Chetouani

Learning from demonstration methods usually leverage close to optimal demonstrations to accelerate training. By contrast, when demonstrating a task, human teachers deviate from optimal demonstrations and pedagogically mo…

Pedagogical Demonstrations and Pragmatic Learning in Artificial Tutor-Learner Interactions

2022-02-28 · Hugo Caselles-Dupré, Mohamed Chetouani, Olivier Sigaud

When demonstrating a task, human tutors pedagogically modify their behavior by either "showing" the task rather than just "doing" it (exaggerating on relevant parts of the demonstration) or by giving demonstrations that …

Emergence of Pragmatics from Referential Game between Theory of Mind Agents

2020-01-21 · Luyao Yuan, Zipeng Fu, Jingyue Shen, Lu Xu 외

Pragmatics studies how context can contribute to language meanings. In human communication, language is never interpreted out of context, and sentences can usually convey more information than their literal meanings. How…

Reinforcement LearningReinforcement Learning (RL)

Pedagogical Alignment of Large Language Models

2024-02-07 · Shashank Sonkar, Kangqi Ni, Sapana Chaudhary, Richard G. Baraniuk

Large Language Models (LLMs), when used in educational settings without pedagogical fine-tuning, often provide immediate answers rather than guiding students through the problem-solving process. This approach falls short…

Synthetic Data Generation

Pedagogical Safety in Educational Reinforcement Learning: Formalizing and Detecting Reward Hacking in AI Tutoring Systems

2026-04-05 · Oluseyi Olukola, Nick Rahimi arxiv

Reinforcement learning (RL) is increasingly used to personalize instruction in intelligent tutoring systems, yet the field lacks a formal framework for defining and evaluating pedagogical safety. We introduce a four-laye…

Reinforcement Learning