paper-with-me

홈 › Papers

A Principle of Targeted Intervention for Multi-Agent Reinforcement Learning

2025-10-20 · Anjie Liu, Jianhong Wang, Samuel Kaski, Jun Wang, Mengyue Yang arxiv

Steering cooperative multi-agent reinforcement learning (MARL) towards desired outcomes is challenging, particularly when the global guidance from a human on the whole multi-agent system is impractical in a large-scale MARL. On the other hand, designing external mechanisms (e.g., intrinsic rewards and human feedback) to coordinate agents mostly relies on empirical studies, lacking a easy-to-use research tool. In this work, we employ multi-agent influence diagrams (MAIDs) as a graphical framework to address the above issues. First, we introduce the concept of MARL interaction paradigms (orthogonal to MARL learning paradigms), using MAIDs to analyze and visualize both unguided self-organization and global guidance mechanisms in MARL. Then, we design a new MARL interaction paradigm, referred to as the targeted intervention paradigm that is applied to only a single targeted agent, so the problem of global guidance can be mitigated. In implementation, we introduce a causal inference technique, referred to as Pre-Strategy Intervention (PSI), to realize the targeted intervention paradigm. Since MAIDs can be regarded as a special class of causal diagrams, a composite desired outcome that integrates the primary task goal and an additional desired outcome can be achieved by maximizing the corresponding causal effect through the PSI. Moreover, the bundled relevance graph analysis of MAIDs provides a tool to identify whether an MARL learning paradigm is workable under the design of an MARL interaction paradigm. In experiments, we demonstrate the effectiveness of our proposed targeted intervention, and verify the result of relevance graph analysis.

📄 PDF Abstract BibTeX arXiv:2510.17697

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement LearningCausal Inference

Similar Papers 제목 키워드 기반

Reinforcement Learning for Electricity Network Operation

2020-03-16 · Adrian Kelly, Aidan O'Sullivan, Patrick de Mars, Antoine Marot

This paper presents the background material required for the Learning to Run Power Networks Challenge. The challenge is focused on using Reinforcement Learning to train an agent to manage the real-time operations of a po…

BIG-bench Machine Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Towards intervention-centric causal reasoning in learning agents

2020-05-26 · Benjamin Lansdell

Interventions are central to causal learning and reasoning. Yet ultimately an intervention is an abstraction: an agent embedded in a physical environment (perhaps modeled as a Markov decision process) does not typically …

Deep Reinforcement LearningMeta-LearningMeta Reinforcement Learningreinforcement-learning+2

Reinforcement Learning Enables Autonomous Microrobot Navigation and Intervention in Simulated Blood Capillaries

2026-06-23 · Jannik Drotleff, Samuel Tovey, Paul Hohenberger, Christoph Lohrmann 외 arxiv

Autonomous microrobots navigating biological vasculature could enable targeted drug delivery and thrombolysis, yet training control policies for realistic environments remains an open challenge. Prior reinforcement learn…

Reinforcement Learning

Adaptive Network Intervention for Complex Systems: A Hierarchical Graph Reinforcement Learning Approach

2024-10-30 · Qiliang Chen, Babak Heydari

Effective governance and steering of behavior in complex multi-agent systems (MAS) are essential for managing system-wide outcomes, particularly in environments where interactions are structured by dynamic networks. In m…

The Containment Gap: How Deployed Agentic AI Frameworks Fail Public-Facing Safety Requirements

2026-06-11 · Md Jafrin Hossain, Mohammad Arif Hossain, Weiqi Liu, Nirwan Ansari arxiv

Agentic large language model systems that autonomously invoke tools, maintain persistent memory, and execute multi-step plans are increasingly deployed in public-facing domains, including government services, healthcare …