paper-with-me

홈 › Papers

HyPOLE: Hyperproperty-Guided Multi-Agent Reinforcement Learning under Partial Observation

2026-06-29 · Arshia Rafieioskouei, Tzu-Han Hsu, Matthew Lucas, Borzoo Bonakdarpour arxiv

Formal specification is a powerful tool to guide the learning process and provides significant advantages over reward shaping: (1) mathematical rigor; (2) expressiveness to specify objectives and constraints, and (3) the ability to define tactics to achieve objectives. However, these benefits remain largely unexplored in the context of Multi-Agent Reinforcement Learning (MARL). This paper introduces HyPOLE, a novel framework for MARL under partial observability, where learning is guided by the expressive power of the so-called hyperproperties and, in particular, the temporal logic HyperLTL. We integrate Centralized Training for Decentralized Execution (CTDE) techniques with HyPOLE to synthesize decentralized policies, and our evaluation on SMAC, MessySMAC, and WildFire benchmark demonstrates clear advantages over baselines.

📄 PDF Abstract BibTeX arXiv:2606.30966

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

Hyperproperty-Constrained Secure Reinforcement Learning

2025-07-31 · Ernest Bonnah, Luan Viet Nguyen, Khaza Anuarul Hoque arxiv

Hyperproperties for Time Window Temporal Logic (HyperTWTL) is a domain-specific formal specification language known for its effectiveness in compactly representing security, opacity, and concurrency properties for roboti…

Reinforcement Learning

On Conformant Planning and Model-Checking of $\exists^*\forall^*$ Hyperproperties

2025-12-29 · Raven Beutner, Bernd Finkbeiner arxiv

We study the connection of two problems within the planning and verification community: Conformant planning and model-checking of hyperproperties. Conformant planning is the task of finding a sequential plan that achieve…

On Alternating-Time Temporal Logic, Hyperproperties, and Strategy Sharing

2023-12-19 · Raven Beutner, Bernd Finkbeiner

Alternating-time temporal logic (ATL$^*$) is a well-established framework for formal reasoning about multi-agent systems. However, while ATL$^*$ can reason about the strategic ability of agents (e.g., some coalition $A$ …

Decentralized Planning Using Probabilistic Hyperproperties

2025-02-19 · Francesco Pontiggia, Filip Macák, Roman Andriushchenko, Michele Chiari 외

Multi-agent planning under stochastic dynamics is usually formalised using decentralized (partially observable) Markov decision processes ( MDPs) and reachability or expected reward specifications. In this paper, we prop…

Non-Deterministic Planning for Hyperproperty Verification

2024-05-22 · Raven Beutner, Bernd Finkbeiner

Non-deterministic planning aims to find a policy that achieves a given objective in an environment where actions have uncertain effects, and the agent - potentially - only observes parts of the current state. Hyperproper…