paper-with-me

Papers

Bayesian adversarial multi-node bandit for optimal smart grid protection against cyber attacks

2021-02-20 · Jianyu Xu, Bin Liu, Huadong Mo, Daoyi Dong

The cybersecurity of smart grids has become one of key problems in developing reliable modern power and energy systems. This paper introduces a non-stationary adversarial cost with a variation constraint for smart grids and enables us to investigate the problem of optimal smart grid protection against cyber attacks in a relatively practical scenario. In particular, a Bayesian multi-node bandit (MNB) model with adversarial costs is constructed and a new regret function is defined for this model. An algorithm called Thompson-Hedge algorithm is presented to solve the problem and the superior performance of the proposed algorithm is proven in terms of the convergence rate of the regret function. The applicability of the algorithm to real smart grid scenarios is verified and the performance of the algorithm is also demonstrated by numerical examples.

📄 PDF Abstract BibTeX arXiv:2104.02774

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Bayesian Design Principles for Frequentist Sequential Learning

2023-10-01 · Yunbei Xu, Assaf Zeevi

We develop a general theory to optimize the frequentist regret for sequential learning problems, where efficient bandit and reinforcement learning algorithms can be derived from unified Bayesian principles. We propose a …

Multi-Armed Banditsreinforcement-learningReinforcement Learning

Continuous-in-time Limit for Bayesian Bandits

2022-10-14 · Yuhua Zhu, Zachary Izzo, Lexing Ying

This paper revisits the bandit problem in the Bayesian setting. The Bayesian approach formulates the bandit problem as an optimization problem, and the goal is to find the optimal policy which minimizes the Bayesian regr…

Pareto Regret Analyses in Multi-objective Multi-armed Bandit

2022-12-01 · Mengfan Xu, Diego Klabjan

We study Pareto optimality in multi-objective multi-armed bandit by providing a formulation of adversarial multi-objective multi-armed bandit and defining its Pareto regrets that can be applied to both stochastic and adv…

Adversarial Attack

Learning-based decentralized offloading decision making in an adversarial environment

2021-04-26 · Byungjin Cho, Yu Xiao

Vehicular fog computing (VFC) pushes the cloud computing capability to the distributed fog nodes at the edge of the Internet, enabling compute-intensive and latency-sensitive computing services for vehicles through task …

Cloud ComputingDecision Making

Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries

2025-04-01 · Arnab Maiti, Zhiyuan Fan, Kevin Jamieson, Lillian J. Ratliff 외

In this paper, we study the online shortest path problem in directed acyclic graphs (DAGs) under bandit feedback against an adaptive adversary. Given a DAG $G = (V, E)$ with a source node $v_{\mathsf{s}}$ and a sink node…

Multi-Armed Bandits