paper-with-me

홈 › Papers

Information-Theoretic Opacity-Enforcement in Markov Decision Processes

2024-04-30 · Chongyang Shi, Yuheng Bu, Jie Fu

The paper studies information-theoretic opacity, an information-flow privacy property, in a setting involving two agents: A planning agent who controls a stochastic system and an observer who partially observes the system states. The goal of the observer is to infer some secret, represented by a random variable, from its partial observations, while the goal of the planning agent is to make the secret maximally opaque to the observer while achieving a satisfactory total return. Modeling the stochastic system using a Markov decision process, two classes of opacity properties are considered -- Last-state opacity is to ensure that the observer is uncertain if the last state is in a specific set and initial-state opacity is to ensure that the observer is unsure of the realization of the initial state. As the measure of opacity, we employ the Shannon conditional entropy capturing the information about the secret revealed by the observable. Then, we develop primal-dual policy gradient methods for opacity-enforcement planning subject to constraints on total returns. We propose novel algorithms to compute the policy gradient of entropy for each observation, leveraging message passing within the hidden Markov models. This gradient computation enables us to have stable and fast convergence. We demonstrate our solution of opacity-enforcement control through a grid world example.

📄 PDF Abstract BibTeX arXiv:2405.00157

Code (0)

등록된 구현이 없습니다.

Tasks

Policy Gradient Methods

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Synthesis of Opacity-Enforcing Winning Strategies Against Colluded Opponent

2023-04-03 · Chongyang Shi, Abhishek N. Kulkarni, Hazhar Rahmani, Jie Fu

This paper studies a language-based opacity enforcement in a two-player, zero-sum game on a graph. In this game, player 1 (P1) wins if it can achieve a secret temporal goal described by the language of a finite automaton…

Motion Planning

Opacity Enforcing Supervisory Control using Non-deterministic Supervisors

2020-10-20 · Yifan Xie, Xiang Yin, ShaoYuan Li

In this paper, we investigate the enforcement of opacity via supervisory control in the context of discrete-event systems. A system is said to be opaque if the intruder, which is modeled as a passive observer, can never …

On Approximate Opacity of Stochastic Control Systems

2024-01-03 · Siyuan Liu, Xiang Yin, Dimos V. Dimarogonas, Majid Zamani

This paper investigates an important class of information-flow security property called opacity for stochastic control systems. Opacity captures whether a system's secret behavior (a subset of the system's behavior that …

Relation

Synthesis of Dynamic Masks for Information-Theoretic Opacity in Stochastic Systems

2025-02-14 · Sumukha Udupa, Chongyang Shi, Jie Fu

In this work, we investigate the synthesis of dynamic information releasing mechanisms, referred to as ''masks'', to minimize information leakage from a stochastic system to an external observer. Specifically, for a stoc…

Privacy-Preserving Co-synthesis Against Sensor-Actuator Eavesdropping Intruder

2021-04-30 · Ruochen Tai, Liyong Lin, Yuting Zhu, Rong Su

In this work, we investigate the problem of privacy-preserving supervisory control against an external passive intruder via co-synthesis of dynamic mask, edit function, and supervisor for opacity enforcement and requirem…

Privacy Preserving