Safe Reinforcement Learning for a Robot Being Pursued but with Objectives Covering More Than Capture-avoidance
Reinforcement Learning (RL) algorithms show amazing performance in recent years, but placing RL in real-world applications such as self-driven vehicles may suffer safety problems. A self-driven vehicle moving to a target position following a learned policy may suffer a vehicle with unpredictable aggressive behaviors or even being pursued by a vehicle following a Nash strategy. To address the safety issue of the self-driven vehicle in this scenario, this paper conducts a preliminary study based on a system of robots. A safe RL framework with safety guarantees is developed for a robot being pursued but with objectives covering more than capture-avoidance. Simulations and experiments are conducted based on the system of robots to evaluate the effectiveness of the developed safe RL framework.
Code (0)
등록된 구현이 없습니다.
Tasks
PositionReinforcement Learning (RL)Safe Reinforcement LearningSimilar Papers 제목 키워드 기반
Dual-Arm Adversarial Robot Learning
Robot learning is a very promising topic for the future of automation and machine intelligence. Future robots should be able to autonomously acquire skills, learn to represent their environment, and interact with it. Whi…
Safe ExplorationRegret Bounds for Risk-Sensitive Reinforcement Learning
In safety-critical applications of reinforcement learning such as healthcare and robotics, it is often desirable to optimize risk-sensitive objectives that account for tail outcomes rather than expected reward. We prove …
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Driver Assistance Eco-driving and Transmission Control with Deep Reinforcement Learning
With the growing need to reduce energy consumption and greenhouse gas emissions, Eco-driving strategies provide a significant opportunity for additional fuel savings on top of other technological solutions being pursued …
Deep Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Automaton Constrained Q-Learning
Real-world robotic tasks often require agents to achieve sequences of goals while respecting time-varying safety constraints. However, standard Reinforcement Learning (RL) paradigms are fundamentally limited in these set…
Reinforcement LearningContinuous ControlSafety and Liveness Guarantees through Reach-Avoid Reinforcement Learning
Reach-avoid optimal control problems, in which the system must reach certain goal conditions while staying clear of unacceptable failure modes, are central to safety and liveness assurance for autonomous robotic systems,…
Deep Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning+1