paper-with-me

홈 › Papers

Counterfactual Effect Decomposition in Multi-Agent Sequential Decision Making

2024-10-16 · Stelios Triantafyllou, Aleksa Sukovic, Yasaman Zolfimoselo, Goran Radanovic

We address the challenge of explaining counterfactual outcomes in multi-agent Markov decision processes. In particular, we aim to explain the total counterfactual effect of an agent's action on the outcome of a realized scenario through its influence on the environment dynamics and the agents' behavior. To achieve this, we introduce a novel causal explanation formula that decomposes the counterfactual effect by attributing to each agent and state variable a score reflecting their respective contributions to the effect. First, we show that the total counterfactual effect of an agent's action can be decomposed into two components: one measuring the effect that propagates through all subsequent agents' actions and another related to the effect that propagates through the state transitions. Building on recent advancements in causal contribution analysis, we further decompose these two effects as follows. For the former, we consider agent-specific effects -- a causal concept that quantifies the counterfactual effect of an agent's action that propagates through a subset of agents. Based on this notion, we use Shapley value to attribute the effect to individual agents. For the latter, we consider the concept of structure-preserving interventions and attribute the effect to state variables based on their "intrinsic" contributions. Through extensive experimentation, we demonstrate the interpretability of our approach in a Gridworld environment with LLM-assisted agents and a sepsis management simulator.

📄 PDF Abstract BibTeX arXiv:2410.12539

Code (1)

stelios30/cf-effect-decomposition 공식 구현

Tasks

AttributecounterfactualDecision MakingSequential Decision Making

Similar Papers 제목 키워드 기반

COSAC: Counterfactual Credit Assignment in Sequential Cooperative Teams

2026-04-20 · Shripad Deshmukh, Jayakumar Subramanian, Raghavendra Addanki, Nikos Vlassis arxiv

In cooperative teams where agents act in a fixed order and share a single team-level reward (multi-agent language systems, sequential robotic tasks), per-agent credit assignment is under-determined. Critic-based approach…

Sequential Counterfactual Decision-Making Under Confounded Reward

2022-06-05 · Erik Skalnes

We investigate the limitations of random trials when the cause of interest is confounded with the effect by formalizing a counterfactual policy-space where the agent's natural predilection is input to a soft-intervention…

counterfactualDecision Making

Towards Understanding Linear Value Decomposition in Cooperative Multi-Agent Q-Learning

2020-09-28 · Jianhao Wang, Zhizhou Ren, Beining Han, Jianing Ye 외

Value decomposition is a popular and promising approach to scaling up multi-agent reinforcement learning in cooperative settings. However, the theoretical understanding of such methods is limited. In this paper, we intro…

counterfactualMulti-agent Reinforcement LearningQ-LearningStarcraft+1

Counterfactual Regularization for Model-Based Reinforcement Learning

2019-09-25 · Lawrence Neal, Li Fuxin, Xiaoli Fern

In sequential tasks, planning-based agents have a number of advantages over model-free agents, including sample efficiency and interpretability. Recurrent action-conditional latent dynamics models trained from pixel-leve…

counterfactualmodelModel-based Reinforcement Learningreinforcement-learning+2

Sequential Transport for Causal Mediation Analysis

2026-03-16 · Agathe Fernandes Machado, Iryna Voitsitska, Arthur Charpentier, Ewen Gallic arxiv

We propose sequential transport (ST), a distributional framework for mediation analysis that combines optimal transport (OT) with a mediator directed acyclic graph (DAG). Instead of relying on cross-world counterfactual …