paper-with-me

Papers

Multi-Agent Decision-Focused Learning via Value-Aware Sequential Communication

2026-04-10 · Benjamin Amoh, Geoffrey Parker, Wesley Marrero arxiv

Multi-agent coordination under partial observability requires agents to share complementary private information. While recent methods optimize messages for intermediate objectives (e.g., reconstruction accuracy or mutual information), rather than decision quality, we introduce \textbf{SeqComm-DFL}, unifying the sequential communication with decision-focused learning for task performance. Our approach features \emph{value-aware message generation with sequential Stackelberg conditioning}: messages maximize receiver decision quality and are generated in priority order, with agents conditioning on their predecessors. The \emph{guidance potential} determined by their prosocial ordering. We extend Optimal Model Design to communication-augmented world models with QMIX factorization, enabling efficient end-to-end training via implicit differentiation. We prove information-theoretic bounds showing that communication value scales with coordination gaps and establish $\mathcal{O}(1/\sqrt{T})$ convergence for the bilevel optimization, where $T$ denotes the number of training iterations. On collaborative healthcare and StarCraft Multi-Agent Challenge (SMAC) benchmarks, SeqComm-DFL achieves four to six times higher cumulative rewards and over 13\% win rate improvements, enabling coordination strategies inaccessible under information asymmetry.

📄 PDF Abstract BibTeX arXiv:2604.08944

Code (0)

등록된 구현이 없습니다.

Tasks

Bilevel Optimization

Similar Papers 제목 키워드 기반

A Scalable Approach to Solving Simulation-Based Network Security Games

2026-02-18 · Michael Lanier, Yevgeniy Vorobeychik arxiv

We introduce MetaDOAR, a lightweight meta-controller that augments the Double Oracle / PSRO paradigm with a learned, partition-aware filtering layer and Q-value caching to enable scalable multi-agent reinforcement learni…

Multi-agent Reinforcement Learning

Explainability in autonomous pedagogically structured scenarios

2022-10-21 · Minal Suresh Patil

We present the notion of explainability for decision-making processes in a pedagogically structured autonomous environment. Multi-agent systems that are structured pedagogically consist of pedagogical teachers and learne…

Decision Making

Agentic LLM Framework for Adaptive Decision Discourse

2025-02-16 · Antoine Dolant, Praveen Kumar

Effective decision-making in complex systems requires synthesizing diverse perspectives to address multifaceted challenges under uncertainty. This study introduces a real-world inspired agentic Large Language Models (LLM…

Navigate

Tradeoff-Focused Contrastive Explanation for MDP Planning

2020-04-27 · Roykrong Sukkerd, Reid Simmons, David Garlan

End-users' trust in automated agents is important as automated decision-making and planning is increasingly used in many aspects of people's lives. In real-world applications of planning, multiple optimization objectives…

Decision MakingRobot Navigation

Risk-Aware Distributed Multi-Agent Reinforcement Learning

2023-04-04 · Abdullah Al Maruf, Luyao Niu, Bhaskar Ramasubramanian, Andrew Clark 외

Autonomous cyber and cyber-physical systems need to perform decision-making, learning, and control in unknown environments. Such decision-making can be sensitive to multiple factors, including modeling errors, changes in…

Decision MakingMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning