paper-with-me

홈 › Papers

Local Differential Privacy for Sequential Decision Making in a Changing Environment

2023-01-02 · Pratik Gajane

We study the problem of preserving privacy while still providing high utility in sequential decision making scenarios in a changing environment. We consider abruptly changing environment: the environment remains constant during periods and it changes at unknown time instants. To formulate this problem, we propose a variant of multi-armed bandits called non-stationary stochastic corrupt bandits. We construct an algorithm called SW-KLUCB-CF and prove an upper bound on its utility using the performance measure of regret. The proven regret upper bound for SW-KLUCB-CF is near-optimal in the number of time steps and matches the best known bound for analogous problems in terms of the number of time steps and the number of changes. Moreover, we present a provably optimal mechanism which can guarantee the desired level of local differential privacy while providing high utility.

📄 PDF Abstract BibTeX arXiv:2301.00561

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingMulti-Armed BanditsSequential Decision Making

Similar Papers 제목 키워드 기반

DP-NCB: Privacy Preserving Fair Bandits

2025-08-05 · Dhruv Sarkar, Nishant Pandey, Sayak Ray Chowdhury arxiv

Multi-armed bandit algorithms are fundamental tools for sequential decision-making under uncertainty, with widespread applications across domains such as clinical trials and personalized decision-making. As bandit algori…

Decision Making in Changing Environments: Robustness, Query-Based Learning, and Differential Privacy

2025-01-24 · Fan Chen, Alexander Rakhlin

We study the problem of interactive decision making in which the underlying environment changes over time subject to given constraints. We propose a framework, which we call \textit{hybrid Decision Making with Structured…

Decision MakingMulti-Armed Bandits

Differentially Private Regret Minimization in Episodic Markov Decision Processes

2021-12-20 · Sayak Ray Chowdhury, Xingyu Zhou

We study regret minimization in finite horizon tabular Markov decision processes (MDPs) under the constraints of differential privacy (DP). This is motivated by the widespread applications of reinforcement learning (RL) …

Decision MakingReinforcement Learning (RL)Sequential Decision Making

Locally Private Nonparametric Contextual Multi-armed Bandits

2025-03-11 · Yuheng Ma, Feiyu Jiang, Zifeng Zhao, Hanfang Yang 외

Motivated by privacy concerns in sequential decision-making on sensitive data, we address the challenge of nonparametric contextual multi-armed bandits (MAB) under local differential privacy (LDP). We develop a uniform-c…

Decision MakingMulti-Armed BanditsSequential Decision Making

Federated Linear Contextual Bandits with User-level Differential Privacy

2023-06-08 · Ruiquan Huang, Huanyu Zhang, Luca Melis, Milan Shen 외

This paper studies federated linear contextual bandits under the notion of user-level differential privacy (DP). We first introduce a unified federated bandits framework that can accommodate various definitions of DP in …

Decision MakingMulti-Armed BanditsSequential Decision Making