paper-with-me

홈 › Papers

Learning Performance Graphs from Demonstrations via Task-Based Evaluations

2022-04-12 · Aniruddh G. Puranic, Jyotirmoy V. Deshmukh, Stefanos Nikolaidis

In the learning from demonstration (LfD) paradigm, understanding and evaluating the demonstrated behaviors plays a critical role in extracting control policies for robots. Without this knowledge, a robot may infer incorrect reward functions that lead to undesirable or unsafe control policies. Recent work has proposed an LfD framework where a user provides a set of formal task specifications to guide LfD, to address the challenge of reward shaping. However, in this framework, specifications are manually ordered in a performance graph (a partial order that specifies relative importance between the specifications). The main contribution of this paper is an algorithm to learn the performance graph directly from the user-provided demonstrations, and show that the reward functions generated using the learned performance graph generate similar policies to those from manually specified performance graphs. We perform a user study that shows that priorities specified by users on behaviors in a simulated highway driving domain match the automatically inferred performance graph. This establishes that we can accurately evaluate user demonstrations with respect to task specifications without expert criteria.

📄 PDF Abstract BibTeX arXiv:2204.05909

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Learning Implicit Causal World Models from Multi-Agent Demonstrations

2026-07-28 · Jasorsi Ghosh arxiv

In model-based reinforcement learning, world models exist as internal simulators, but their training often conflates statistical correlations with causal mechanisms. This problem is exacerbated in multi-agent systems whe…

Reinforcement Learning

TaPeR: Probabilistic Recovery of Sparse Task Precedence Graphs from a Handful of Demonstrations

2026-08-21 · Adrian Röfer, Karla Stepanova, Abhinav Valada arxiv

Long-horizon manipulation tasks are often only partially ordered. For example, when assembling an electronic device, the battery and circuit board may be installed in either order, but both must be in place before the en…

Compositional Servoing by Recombining Demonstrations

2023-10-06 · Max Argus, Abhijeet Nayak, Martin Büchner, Silvio Galesso 외

Learning-based manipulation policies from image inputs often show weak task transfer capabilities. In contrast, visual servoing methods allow efficient task transfer in high-precision scenarios while requiring only a few…

Mixture of Demonstrations for Textual Graph Understanding and Question Answering

2026-03-23 · Yukun Wu, Lihui Liu arxiv

Textual graph-based retrieval-augmented generation (GraphRAG) has emerged as a powerful paradigm for enhancing large language models (LLMs) in domain-specific question answering. While existing approaches primarily focus…

Question Answering

Learning API Functionality from Demonstrations for Tool-based Agents

2025-05-30 · Bhrij Patel, Ashish Jagmohan, Aditya Vempaty

Digital tool-based agents that invoke external Application Programming Interfaces (APIs) often rely on documentation to understand API functionality. However, such documentation is frequently missing, outdated, privatize…