paper-with-me

홈 › Papers

Learning to Assist Humans without Inferring Rewards

2024-11-04 · Vivek Myers, Evan Ellis, Sergey Levine, Benjamin Eysenbach, Anca Dragan

Assistive agents should make humans' lives easier. Classically, such assistance is studied through the lens of inverse reinforcement learning, where an assistive agent (e.g., a chatbot, a robot) infers a human's intention and then selects actions to help the human reach that goal. This approach requires inferring intentions, which can be difficult in high-dimensional settings. We build upon prior work that studies assistance through the lens of empowerment: an assistive agent aims to maximize the influence of the human's actions such that they exert a greater control over the environmental outcomes and can solve tasks in fewer steps. We lift the major limitation of prior work in this area--scalability to high-dimensional settings--with contrastive successor representations. We formally prove that these representations estimate a similar notion of empowerment to that studied by prior work and provide a ready-made mechanism for optimizing it. Empirically, our proposed method outperforms prior methods on synthetic benchmarks, and scales to Overcooked, a cooperative game setting. Theoretically, our work connects ideas from information theory, neuroscience, and reinforcement learning, and charts a path for representations to play a critical role in solving assistive problems.

📄 PDF Abstract BibTeX arXiv:2411.02623

Code (1)

vivekmyers/empowerment_successor_representations 공식 구현 jax

Tasks

Chatbotreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Training LLM Agents to Empower Humans

2025-10-15 · Evan Ellis, Vivek Myers, Jens Tuyls, Sergey Levine 외 arxiv

Assistive agents should not only take actions on behalf of a human, but also step out of the way and cede control when there are important decisions to be made. However, current methods for building assistive agents, whe…

Emergent social transmission of model-based representations without inference

2026-04-07 · Silja Keßler, Miriam Bautista-Salinero, Claudio Tennie, Charley M. Wu arxiv

How do people acquire rich, flexible knowledge about their environment from others despite limited cognitive capacity? Humans are often thought to rely on computationally costly mentalizing, such as inferring others' bel…

Reinforcement Learning

The Assistive Multi-Armed Bandit

2019-01-24 · Lawrence Chan, Dylan Hadfield-Menell, Siddhartha Srinivasa, Anca Dragan

Learning preferences implicit in the choices humans make is a well studied problem in both economics and computer science. However, most work makes the assumption that humans are acting (noisily) optimally with respect t…

Multi-Armed Bandits

COOPERA: Continual Open-Ended Human-Robot Assistance

2025-10-27 · Chenyang Ma, Kai Lu, Ruta Desai, Xavier Puig 외 arxiv

To understand and collaborate with humans, robots must account for individual human traits, habits, and activities over time. However, most robotic assistants lack these abilities, as they primarily focus on predefined t…

Inferring Personalized Bayesian Embeddings for Learning from Heterogeneous Demonstration

2019-03-14 · Rohan Paleja, Matthew Gombolay

For assistive robots and virtual agents to achieve ubiquity, machines will need to anticipate the needs of their human counterparts. The field of Learning from Demonstration (LfD) has sought to enable machines to infer p…

Decision Making