paper-with-me

Papers

Byzantine-Robust Online and Offline Distributed Reinforcement Learning

2022-06-01 · Yiding Chen, Xuezhou Zhang, Kaiqing Zhang, Mengdi Wang, Xiaojin Zhu

We consider a distributed reinforcement learning setting where multiple agents separately explore the environment and communicate their experiences through a central server. However, $\alpha$-fraction of agents are adversarial and can report arbitrary fake information. Critically, these adversarial agents can collude and their fake data can be of any sizes. We desire to robustly identify a near-optimal policy for the underlying Markov decision process in the presence of these adversarial agents. Our main technical contribution is Weighted-Clique, a novel algorithm for the robust mean estimation from batches problem, that can handle arbitrary batch sizes. Building upon this new estimator, in the offline setting, we design a Byzantine-robust distributed pessimistic value iteration algorithm; in the online setting, we design a Byzantine-robust distributed optimistic value iteration algorithm. Both algorithms obtain near-optimal sample complexities and achieve superior robustness guarantee than prior works.

📄 PDF Abstract BibTeX arXiv:2206.00165

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Byzantine-Robust Distributed Online Learning: Taming Adversarial Participants in An Adversarial Environment

2023-07-16 · Xingrong Dong, Zhaoxian Wu, Qing Ling, Zhi Tian

This paper studies distributed online learning under Byzantine attacks. The performance of an online learning algorithm is often characterized by (adversarial) regret, which evaluates the quality of one-step-ahead decisi…

Decision Making

Distributed Online Optimization with Byzantine Adversarial Agents

2021-09-25 · Sourav Sahoo, Anand Gokhale, Rachel Kalpana Kalaimani

We study the problem of non-constrained, discrete-time, online distributed optimization in a multi-agent system where some of the agents do not follow the prescribed update rule either due to failures or malicious intent…

Distributed Optimization

Secure Byzantine-Robust Distributed Learning via Clustering

2021-10-06 · Raj Kiriti Velicheti, Derek Xia, Oluwasanmi Koyejo

Federated learning systems that jointly preserve Byzantine robustness and privacy have remained an open problem. Robust aggregation, the standard defense for Byzantine attacks, generally requires server access to individ…

ClusteringFederated LearningPrivacy Preserving

Online Multi-Agent Decentralized Byzantine-robust Gradient Estimation

2022-09-30 · Alexandre Reiffers-Masson, Isabel Amigo

In this paper, we propose an iterative scheme for distributed Byzantineresilient estimation of a gradient associated with a black-box model. Our algorithm is based on simultaneous perturbation, secure state estimation an…

State Estimation

Byzantine-Resilient Distributed P2P Energy Trading via Spatial-Temporal Anomaly Detection

2025-05-26 · Junhong Liu, Qinfei Long, Rong-Peng Liu, Wenjie Liu 외

Distributed peer-to-peer (P2P) energy trading mandates an escalating coupling between the physical power network and communication network, necessitating high-frequency sharing of real-time data among prosumers. However,…

Anomaly DetectionComputational Efficiencyenergy trading