paper-with-me

홈 › Papers

A Framework for Sequential Planning in Multi-Agent Settings

2011-09-09 · P. Doshi, P. J. Gmytrasiewicz

This paper extends the framework of partially observable Markov decision processes (POMDPs) to multi-agent settings by incorporating the notion of agent models into the state space. Agents maintain beliefs over physical states of the environment and over models of other agents, and they use Bayesian updates to maintain their beliefs over time. The solutions map belief states to actions. Models of other agents may include their belief states and are related to agent types considered in games of incomplete information. We express the agents autonomy by postulating that their models are not directly manipulable or observable by other agents. We show that important properties of POMDPs, such as convergence of value iteration, the rate of convergence, and piece-wise linearity and convexity of the value functions carry over to our framework. Our approach complements a more traditional approach to interactive settings which uses Nash equilibria as a solution paradigm. We seek to avoid some of the drawbacks of equilibria which may be non-unique and do not capture off-equilibrium behaviors. We do so at the cost of having to represent, process and continuously revise models of other agents. Since the agents beliefs may be arbitrarily nested, the optimal solutions to decision making problems are only asymptotically computable. However, approximate belief updates and approximately optimal plans are computable. We illustrate our framework using a simple application domain, and we show examples of belief updates and value functions.

📄 PDF Abstract BibTeX arXiv:1109.2135

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Learning Others' Intentional Models in Multi-Agent Settings Using Interactive POMDPs

2018-12-01 · NeurIPS 2018 12 · Yanlin Han, Piotr Gmytrasiewicz

Interactive partially observable Markov decision processes (I-POMDPs) provide a principled framework for planning and acting in a partially observable, stochastic and multi-agent environment. It extends POMDPs to multi-a…

Bayesian Inference

Cooperative Epistemic Multi-Agent Planning for Implicit Coordination

2017-03-07 · Thorsten Engesser, Thomas Bolander, Robert Mattmüller, Bernhard Nebel

Epistemic planning can be used for decision making in multi-agent situations with distributed knowledge and capabilities. Recently, Dynamic Epistemic Logic (DEL) has been shown to provide a very natural and expressive fr…

Decision Making

HiMAP-Travel: Hierarchical Multi-Agent Planning for Long-Horizon Constrained Travel

2026-03-05 · The Viet Bui, Wenjun Li, Yong Liu arxiv

Sequential LLM agents fail on long-horizon planning with hard constraints like budgets and diversity requirements. As planning progresses and context grows, these agents drift from global constraints. We propose HiMAP-Tr…

Enhancing Temporal Planning Domains by Sequential Macro-actions (Extended Version)

2023-07-22 · Marco De Bortoli, Lukáš Chrpa, Martin Gebser, Gerald Steinbauer-Wagner

Temporal planning is an extension of classical planning involving concurrent execution of actions and alignment with temporal constraints. Durative actions along with invariants allow for modeling domains in which multip…

Path Planning for a Cooperative Navigation Aid Vehicle to Assist Multiple Agents Sequentially

2024-02-26 · Artur Wolek

This paper considers planning a path for a single underwater cooperative navigation aid (CNA) vehicle to sequentially aid a set of N agents to minimize average navigation uncertainty. Both the CNA and agents are modeled …