Causal Responsibility Attribution for Human-AI Collaboration
As Artificial Intelligence (AI) systems increasingly influence decision-making across various fields, the need to attribute responsibility for undesirable outcomes has become essential, though complicated by the complex interplay between humans and AI. Existing attribution methods based on actual causality and Shapley values tend to disproportionately blame agents who contribute more to an outcome and rely on real-world measures of blameworthiness that may misalign with responsible AI standards. This paper presents a causal framework using Structural Causal Models (SCMs) to systematically attribute responsibility in human-AI systems, measuring overall blameworthiness while employing counterfactual reasoning to account for agents' expected epistemic levels. Two case studies illustrate the framework's adaptability in diverse human-AI collaboration scenarios.
Code (1)
Tasks
AttributecounterfactualCounterfactual ReasoningDecision MakingSimilar Papers 제목 키워드 기반
Actual Causality and Responsibility Attribution in Decentralized Partially Observable Markov Decision Processes
Actual causality and a closely related concept of responsibility attribution are central to accountable decision making. Actual causality focuses on specific outcomes and aims to identify decisions (actions) that were cr…
Decision MakingDecision Making Under UncertaintySequential Decision MakingTowards Computationally Efficient Responsibility Attribution in Decentralized Partially Observable MDPs
Responsibility attribution is a key concept of accountable multi-agent decision making. Given a sequence of actions, responsibility attribution mechanisms quantify the impact of each participating agent to the final outc…
Card GamesDecision MakingHuman Attribution of Causality to AI Across Agency, Misuse, and Misalignment
AI-related incidents are becoming increasingly frequent and severe, ranging from safety failures to misuse by malicious actors. In such complex situations, identifying which elements caused an adverse outcome, the proble…
Biased Error Attribution in Multi-Agent Human-AI Systems Under Delayed Feedback
Human decision-making is strongly influenced by cognitive biases, particularly under conditions of uncertainty and risk. While prior work has examined bias in single-step decisions with immediate outcomes and in human in…
Counterfactual Reasoning for Causal Responsibility Attribution in Probabilistic Multi-Agent Systems
Responsibility allocation -- determining the extent to which agents are accountable for outcomes -- is a fundamental challenge in the design and analysis of multi-agent systems. In this work, we model such systems as con…