Deceptive Reinforcement Learning for Privacy-Preserving Planning
In this paper, we study the problem of deceptive reinforcement learning to preserve the privacy of a reward function. Reinforcement learning is the problem of finding a behaviour policy based on rewards received from exploratory behaviour. A key ingredient in reinforcement learning is a reward function, which determines how much reward (negative or positive) is given and when. However, in some situations, we may want to keep a reward function private; that is, to make it difficult for an observer to determine the reward function used. We define the problem of privacy-preserving reinforcement learning, and present two models for solving it. These models are based on dissimulation -- a form of deception that `hides the truth'. We evaluate our models both computationally and via human behavioural experiments. Results show that the resulting policies are indeed deceptive, and that participants can determine the true reward function less reliably than that of an honest agent.
Code (0)
등록된 구현이 없습니다.
Tasks
Privacy Preservingreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Deceptive Reinforcement Learning in Model-Free Domains
This paper investigates deceptive reinforcement learning for privacy preservation in model-free and continuous action space domains. In reinforcement learning, the reward function defines the agent's objective. In advers…
modelreinforcement-learningReinforcement LearningReinforcement Learning (RL)Superstition in the Network: Deep Reinforcement Learning Plays Deceptive Games
Deep reinforcement learning has learned to play many games well, but failed on others. To better characterize the modes and reasons of failure of deep reinforcement learners, we test the widely used Asynchronous Actor-Cr…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Repeated Deceptive Path Planning against Learnable Observer
We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work assumes static, non-learning observers, real-world adversaries-such …
Go-Explore for Residential Energy Management
Reinforcement learning is commonly applied in residential energy management, particularly for optimizing energy costs. However, RL agents often face challenges when dealing with deceptive and sparse rewards in the energy…
Efficient Explorationenergy managementManagementreinforcement-learning+1Locally Differentially Private Reinforcement Learning for Linear Mixture Markov Decision Processes
Reinforcement learning (RL) algorithms can be used to provide personalized services, which rely on users' private and sensitive data. To protect the users' privacy, privacy-preserving RL algorithms are in demand. In this…
Privacy Preservingreinforcement-learningReinforcement Learning (RL)