paper-with-me

홈 › Papers

Fast Lifelong Adaptive Inverse Reinforcement Learning from Demonstrations

2022-09-24 · Letian Chen, Sravan Jayanthi, Rohan Paleja, Daniel Martin, Viacheslav Zakharov, Matthew Gombolay

Learning from Demonstration (LfD) approaches empower end-users to teach robots novel tasks via demonstrations of the desired behaviors, democratizing access to robotics. However, current LfD frameworks are not capable of fast adaptation to heterogeneous human demonstrations nor the large-scale deployment in ubiquitous robotics applications. In this paper, we propose a novel LfD framework, Fast Lifelong Adaptive Inverse Reinforcement learning (FLAIR). Our approach (1) leverages learned strategies to construct policy mixtures for fast adaptation to new demonstrations, allowing for quick end-user personalization, (2) distills common knowledge across demonstrations, achieving accurate task inference; and (3) expands its model only when needed in lifelong deployments, maintaining a concise set of prototypical strategies that can approximate all behaviors via policy mixtures. We empirically validate that FLAIR achieves adaptability (i.e., the robot adapts to heterogeneous, user-specific task preferences), efficiency (i.e., the robot achieves sample-efficient adaptation), and scalability (i.e., the model grows sublinearly with the number of demonstrations while maintaining high performance). FLAIR surpasses benchmarks across three control tasks with an average 57% improvement in policy returns and an average 78% fewer episodes required for demonstration modeling using policy mixtures. Finally, we demonstrate the success of FLAIR in a table tennis task and find users rate FLAIR as having higher task (p<.05) and personalization (p<.05) performance.

📄 PDF Abstract BibTeX arXiv:2209.11908

Code (0)

등록된 구현이 없습니다.

Tasks

Continuous Controlreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Lifelong Inverse Reinforcement Learning

2022-07-01 · NeurIPS 2018 12 · Jorge A. Mendez, Shashank Shivkumar, Eric Eaton

Methods for learning from demonstration (LfD) have shown success in acquiring behavior policies by imitating a user. However, even for a single task, LfD may require numerous demonstrations. For versatile agents that mus…

Lifelong learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Environment Design for Inverse Reinforcement Learning

2022-10-26 · Thomas Kleine Buening, Victor Villin, Christos Dimitrakakis

Learning a reward function from demonstrations suffers from low sample-efficiency. Even with abundant data, current inverse reinforcement learning methods that focus on learning from a single environment can fail to hand…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

X-MEN: Guaranteed XOR-Maximum Entropy Constrained Inverse Reinforcement Learning

2022-03-22 · Fan Ding, Yeiang Xue

Inverse Reinforcement Learning (IRL) is a powerful way of learning from demonstrations. In this paper, we address IRL problems with the availability of prior knowledge that optimal policies will never violate certain con…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Multi-Task Lifelong Reinforcement Learning for Wireless Sensor Networks

2025-06-19 · Hossein Mohammadi Firouzjaei, Rafaela Scaciota, Sumudu Samarakoon

Enhancing the sustainability and efficiency of wireless sensor networks (WSN) in dynamic and unpredictable environments requires adaptive communication and energy harvesting strategies. We propose a novel adaptive contro…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Deep Adaptive Multi-Intention Inverse Reinforcement Learning

2021-07-14 · Ariyan Bighashdel, Panagiotis Meletis, Pavol Jancura, Gijs Dubbelman

This paper presents a deep Inverse Reinforcement Learning (IRL) framework that can learn an a priori unknown number of nonlinear reward functions from unlabeled experts' demonstrations. For this purpose, we employ the to…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)