Learning to Assist Agents by Observing Them
The ability of an AI agent to assist other agents, such as humans, is an important and challenging goal, which requires the assisting agent to reason about the behavior and infer the goals of the assisted agent. Training such an ability by using reinforcement learning usually requires large amounts of online training, which is difficult and costly. On the other hand, offline data about the behavior of the assisted agent might be available, but is non-trivial to take advantage of by methods such as offline reinforcement learning. We introduce methods where the capability to create a representation of the behavior is first pre-trained with offline data, after which only a small amount of interaction data is needed to learn an assisting policy. We test the setting in a gridworld where the helper agent has the capability to manipulate the environment of the assisted artificial agents, and introduce three different scenarios where the assistance considerably improves the performance of the assisted agents.
Code (0)
등록된 구현이 없습니다.
Tasks
AI Agentreinforcement-learningReinforcement LearningReinforcement Learning (RL)Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
HoloAssist: an Egocentric Human Interaction Dataset for Interactive AI Assistants in the Real World
Building an interactive AI assistant that can perceive, reason, and collaborate with humans in the real world has been a long-standing pursuit in the AI community. This work is part of a broader research effort to develo…
Mistake DetectionMixed RealityType predictionRobot Learning Theory of Mind through Self-Observation: Exploiting the Intentions-Beliefs Synergy
In complex environments, where the human sensory system reaches its limits, our behaviour is strongly driven by our beliefs about the state of the world around us. Accessing others' beliefs, intentions, or mental states …
AttributeLearning TheoryDomain-independent generation and classification of behavior traces
Financial institutions mostly deal with people. Therefore, characterizing different kinds of human behavior can greatly help institutions for improving their relation with customers and with regulatory offices. In many o…
ClassificationGeneral ClassificationLearning Actionable Representations from Visual Observations
In this work we explore a new approach for robots to teach themselves about the world simply by observing it. In particular we investigate the effectiveness of learning task-agnostic representations for continuous contro…
continuous-controlContinuous ControlReinforcement LearningLLM experiments with simulation: Large Language Model Multi-Agent System for Simulation Model Parametrization in Digital Twins
This paper presents a novel design of a multi-agent system framework that applies large language models (LLMs) to automate the parametrization of simulation models in digital twins. This framework features specialized LL…
Decision MakingLanguage ModelingLanguage ModellingLarge Language Model+1