paper-with-me

Papers

Density Constrained Reinforcement Learning

2021-06-24 · Zengyi Qin, Yuxiao Chen, Chuchu Fan

We study constrained reinforcement learning (CRL) from a novel perspective by setting constraints directly on state density functions, rather than the value functions considered by previous works. State density has a clear physical and mathematical interpretation, and is able to express a wide variety of constraints such as resource limits and safety requirements. Density constraints can also avoid the time-consuming process of designing and tuning cost functions required by value function-based constraints to encode system specifications. We leverage the duality between density functions and Q functions to develop an effective algorithm to solve the density constrained RL problem optimally and the constrains are guaranteed to be satisfied. We prove that the proposed algorithm converges to a near-optimal solution with a bounded error even when the policy update is imperfect. We use a set of comprehensive experiments to demonstrate the advantages of our approach over state-of-the-art CRL methods, with a wide range of density constrained tasks as well as standard CRL benchmarks such as Safety-Gym.

📄 PDF Abstract BibTeX arXiv:2106.12764

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Constrained Policy Optimization with Explicit Behavior Density for Offline Reinforcement Learning

2023-01-28 · NeurIPS 2023 11 · Jing Zhang, Chi Zhang, Wenjia Wang, Bing-Yi Jing

Due to the inability to interact with the environment, offline reinforcement learning (RL) methods face the challenge of estimating the Out-of-Distribution (OOD) points. Existing methods for addressing this issue either …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Supported Trust Region Optimization for Offline Reinforcement Learning

2023-11-15 · Yixiu Mao, Hongchang Zhang, Chen Chen, Yi Xu 외

Offline reinforcement learning suffers from the out-of-distribution issue and extrapolation error. Most policy constraint methods regularize the density of the trained policy towards the behavior policy, which is too res…

MuJoCoreinforcement-learningReinforcement Learning

Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement Learning

2024-12-17 · Chenglin Li, Guangchun Ruan, Hua Geng

Safe reinforcement learning (RL) is a popular and versatile paradigm to learn reward-maximizing policies with safety guarantees. Previous works tend to express the safety constraints in an expectation form due to the eas…

Formreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Density-Aware Reinforcement Learning to Optimise Energy Efficiency in UAV-Assisted Networks

2023-06-14 · Babatunji Omoniwa, Boris Galkin, Ivana Dusparic

Unmanned aerial vehicles (UAVs) serving as aerial base stations can be deployed to provide wireless connectivity to mobile users, such as vehicles. However, the density of vehicles on roads often varies spatially and tem…

Multi-agent Reinforcement Learningreinforcement-learning

Linear Convergence of the Subspace Constrained Mean Shift Algorithm: From Euclidean to Directional Data

2021-04-29 · Yikun Zhang, Yen-Chi Chen

This paper studies the linear convergence of the subspace constrained mean shift (SCMS) algorithm, a well-known algorithm for identifying a density ridge defined by a kernel density estimator. By arguing that the SCMS al…