paper-with-me

Papers

How to talk so AI will learn: Instructions, descriptions, and autonomy

2022-06-16 · Theodore R Sumers, Robert D Hawkins, Mark K Ho, Thomas L Griffiths, Dylan Hadfield-Menell

From the earliest years of our lives, humans use language to express our beliefs and desires. Being able to talk to artificial agents about our preferences would thus fulfill a central goal of value alignment. Yet today, we lack computational models explaining such language use. To address this challenge, we formalize learning from language in a contextual bandit setting and ask how a human might communicate preferences over behaviors. We study two distinct types of language: $\textit{instructions}$, which provide information about the desired policy, and $\textit{descriptions}$, which provide information about the reward function. We show that the agent's degree of autonomy determines which form of language is optimal: instructions are better in low-autonomy settings, but descriptions are better when the agent will need to act independently. We then define a pragmatic listener agent that robustly infers the speaker's reward function by reasoning about $\textit{how}$ the speaker expresses themselves. We validate our models with a behavioral experiment, demonstrating that (1) our speaker model predicts human behavior, and (2) our pragmatic listener successfully recovers humans' reward functions. Finally, we show that this form of social learning can integrate with and reduce regret in traditional reinforcement learning. We hope these insights facilitate a shift from developing agents that $\textit{obey}$ language to agents that $\textit{learn}$ from it.

📄 PDF Abstract BibTeX arXiv:2206.07870

Code (1)

tsumers/how-to-talk 공식 구현

Similar Papers 제목 키워드 기반

AURA: Multimodal Shared Autonomy for Real-World Urban Navigation

2026-04-02 · Yukai Ma, Honglin He, Selina Song, Wayne Wu 외 arxiv

Long-horizon navigation in complex urban environments relies heavily on continuous human operation, which leads to fatigue, reduced efficiency, and safety concerns. Shared autonomy, where a Vision-Language AI agent and a…

Generating Instructions at Different Levels of Abstraction

2020-10-08 · COLING 2020 8 · Arne Köhn, Julia Wichlacz, Álvaro Torralba, Daniel Höller 외

When generating technical instructions, it is often convenient to describe complex objects in the world at different levels of abstraction. A novice user might need an object explained piece by piece, while for an expert…

MinecraftObject

Autonomy and Safety Assurance in the Early Development of Robotics and Autonomous Systems

2025-01-30 · Dhaminda B. Abeywickrama, Michael Fisher, Frederic Wheeler, Louise Dennis

This report provides an overview of the workshop titled Autonomy and Safety Assurance in the Early Development of Robotics and Autonomous Systems, hosted by the Centre for Robotic Autonomy in Demanding and Long-Lasting E…

TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

2023-04-01 · Yifeng Ma, Suzhen Wang, Yu Ding, Bowen Ma 외

Audio-driven talking head generation has drawn growing attention. To produce talking head videos with desired facial expressions, previous methods rely on extra reference videos to provide expression information, which m…

2D Semantic Segmentation task 3 (25 classes)Talking Head Generation

Exploring how EFL students talk to and through AI to develop texts

2026-04-06 · David James Woo, Yangyang Yu, Yilin Huang, Deliang Wang 외 arxiv

Generative Artificial Intelligence (AI) introduces new considerations for English as a foreign language (EFL) writing pedagogy. This study explores how students talk to and through AI by prompt engineering and negotiatin…

Prompt Engineering