paper-with-me

홈 › Papers

Imitating careful experts to avoid catastrophic events

2023-02-02 · Jack R. P. Hanslope, Laurence Aitchison

RL is increasingly being used to control robotic systems that interact closely with humans. This interaction raises the problem of safe RL: how to ensure that a RL-controlled robotic system never, for instance, injures a human. This problem is especially challenging in rich, realistic settings where it is not even possible to clearly write down a reward function which incorporates these outcomes. In these circumstances, perhaps the only viable approach is based on IRL, which infers rewards from human demonstrations. However, IRL is massively underdetermined as many different rewards can lead to the same optimal policies; we show that this makes it difficult to distinguish catastrophic outcomes (such as injuring a human) from merely undesirable outcomes. Our key insight is that humans do display different behaviour when catastrophic outcomes are possible: they become much more careful. We incorporate carefulness signals into IRL, and find that they do indeed allow IRL to disambiguate undesirable from catastrophic outcomes, which is critical to ensuring safety in future real-world human-robot interactions.

📄 PDF Abstract BibTeX arXiv:2302.01193

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Robust Imitation Learning against Variations in Environment Dynamics

2022-06-19 · Jongseong Chae, Seungyul Han, Whiyoung Jung, Myungsik Cho 외

In this paper, we propose a robust imitation learning (IL) framework that improves the robustness of IL when environment dynamics are perturbed. The existing IL framework trained in a single environment can catastrophica…

Imitation Learning

Routing Networks with Co-training for Continual Learning

2020-09-09 · Mark Collier, Efi Kokiopoulou, Andrea Gesmundo, Jesse Berent

The core challenge with continual learning is catastrophic forgetting, the phenomenon that when neural networks are trained on a sequence of tasks they rapidly forget previously learned tasks. It has been observed that c…

Continual Learning

A Taxonomy of Omnicidal Futures Involving Artificial Intelligence

2025-07-12 · Andrew Critch, Jacob Tsimerman arxiv

This report presents a taxonomy and examples of potential omnicidal events resulting from AI: scenarios where all or almost all humans are killed. These events are not presented as inevitable, but as possibilities that w…

Synaptic metaplasticity in binarized neural networks

2021-01-19 · Axel Laborieux, Maxence Ernoult, Tifenn Hirtzlin, Damien Querlioz

Unlike the brain, artificial neural networks, including state-of-the-art deep neural networks for computer vision, are subject to "catastrophic forgetting": they rapidly forget the previous task when trained on a new one…

RAIL: Risk-Averse Imitation Learning

2017-07-20 · Anirban Santara, Abhishek Naik, Balaraman Ravindran, Dipankar Das 외

Imitation learning algorithms learn viable policies by imitating an expert's behavior when reward signals are not available. Generative Adversarial Imitation Learning (GAIL) is a state-of-the-art algorithm for learning p…

Autonomous Drivingcontinuous-controlContinuous ControlImitation Learning