paper-with-me

Papers

Generalized Inverse Planning: Learning Lifted non-Markovian Utility for Generalizable Task Representation

2020-11-12 · Sirui Xie, Feng Gao, Song-Chun Zhu

In searching for a generalizable representation of temporally extended tasks, we spot two necessary constituents: the utility needs to be non-Markovian to transfer temporal relations invariant to a probability shift, the utility also needs to be lifted to abstract out specific grounding objects. In this work, we study learning such utility from human demonstrations. While inverse reinforcement learning (IRL) has been accepted as a general framework of utility learning, its fundamental formulation is one concrete Markov Decision Process. Thus the learned reward function does not specify the task independently of the environment. Going beyond that, we define a domain of generalization that spans a set of planning problems following a schema. We hence propose a new quest, Generalized Inverse Planning, for utility learning in this domain. We further outline a computational framework, Maximum Entropy Inverse Planning (MEIP), that learns non-Markovian utility and associated concepts in a generative manner. The learned utility and concepts form a task representation that generalizes regardless of probability shift or structural change. Seeing that the proposed generalization problem has not been widely studied yet, we carefully define an evaluation protocol, with which we illustrate the effectiveness of MEIP on two proof-of-concept domains and one challenging task: learning to fold from demonstrations.

📄 PDF Abstract BibTeX arXiv:2011.09854

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Novelty and Lifted Helpful Actions in Generalized Planning

2023-07-03 · Chao Lei, Nir Lipovetzky, Krista A. Ehinger

It has been shown recently that successful techniques in classical planning, such as goal-oriented heuristics and landmarks, can improve the ability to compute planning programs for generalized planning (GP) problems. In…

Stochastic Planning and Lifted Inference

2017-01-04 · Roni Khardon, Scott Sanner

Lifted probabilistic inference (Poole, 2003) and symbolic dynamic programming for lifted stochastic planning (Boutilier et al, 2001) were introduced around the same time as algorithmic efforts to use abstraction in stoch…

Decision MakingSequential Decision Making

PG3: Policy-Guided Planning for Generalized Policy Generation

2022-04-21 · Ryan Yang, Tom Silver, Aidan Curtis, Tomas Lozano-Perez 외

A longstanding objective in classical planning is to synthesize policies that generalize across multiple problems from the same domain. In this work, we study generalized policy search-based methods with a focus on the s…

Hoeffding's Inequality for Markov Chains under Generalized Concentrability Condition

2023-10-04 · Hao Chen, Abhishek Gupta, Yin Sun, Ness Shroff

This paper studies Hoeffding's inequality for Markov chains under the generalized concentrability condition defined via integral probability metric (IPM). The generalized concentrability condition establishes a framework…

Generalized Planning: Non-Deterministic Abstractions and Trajectory Constraints

2019-09-26 · Blai Bonet, Giuseppe De Giacomo, Hector Geffner, Sasha Rubin

We study the characterization and computation of general policies for families of problems that share a structure characterized by a common reduction into a single abstract problem. Policies $\mu$ that solve the abstract…