paper-with-me

Papers

The Accountability Horizon: An Impossibility Theorem for Governing Human-Agent Collectives

2026-04-09 · Haileleol Tibebu, Hewan Shemtaga arxiv

Existing accountability frameworks for AI systems, legal, ethical, and regulatory, rest on a shared assumption: for any consequential outcome, at least one identifiable person had enough involvement and foresight to bear meaningful responsibility. This paper proves that agentic AI systems violate this assumption not as an engineering limitation but as a mathematical necessity once autonomy exceeds a computable threshold. We introduce Human-Agent Collectives, a formalisation of joint human-AI systems where agents are modelled as state-policy tuples within a shared structural causal model. Autonomy is characterised through a four-dimensional information-theoretic profile (epistemic, executive, evaluative, social); collective behaviour through interaction graphs and joint action spaces. We axiomatise legitimate accountability through four minimal properties: Attributability (responsibility requires causal contribution), Foreseeability Bound (responsibility cannot exceed predictive capacity), Non-Vacuity (at least one agent bears non-trivial responsibility), and Completeness (all responsibility must be fully allocated). Our central result, the Accountability Incompleteness Theorem, proves that for any collective whose compound autonomy exceeds the Accountability Horizon and whose interaction graph contains a human-AI feedback cycle, no framework can satisfy all four properties simultaneously. The impossibility is structural: transparency, audits, and oversight cannot resolve it without reducing autonomy. Below the threshold, legitimate frameworks exist, establishing a sharp phase transition. Experiments on 3,000 synthetic collectives confirm all predictions with zero violations. This is the first impossibility result in AI governance, establishing a formal boundary below which current paradigms remain valid and above which distributed accountability mechanisms become necessary.

📄 PDF Abstract BibTeX arXiv:2604.07778

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Information-Theoretic Limits of Safety Verification for Self-Improving Systems

2026-03-30 · Arsenios Scrivens arxiv

Can a safety gate permit unbounded beneficial self-modification while maintaining bounded cumulative risk? We formalize this question through dual conditions -- requiring sum delta_n < infinity (bounded risk) and sum TPR…

Automated Search for Impossibility Theorems in Social Choice Theory: Ranking Sets of Objects

2014-01-16 · Christian Geist, Ulle Endriss

We present a method for using standard techniques from satisfiability checking to automatically verify and discover theorems in an area of economic theory known as ranking sets of objects. The key question in this area, …

Decision MakingDecision Making Under Uncertainty

Impossibility and Uncertainty Theorems in AI Value Alignment (or why your AGI should not have a utility function)

2018-12-31 · Peter Eckersley

Utility functions or their equivalents (value functions, objective functions, loss functions, reward functions, preference orderings) are a central tool in most current machine learning systems. These mechanisms for defi…

BIG-bench Machine Learning

Aggregating Credences into Beliefs: Agenda Conditions for Impossibility Results

2023-07-11 · Minkyung Wang, Chisu Kim

Binarizing belief aggregation addresses how to rationally aggregate individual probabilistic beliefs into collective binary beliefs. Similar to the development of judgment aggregation theory, formulating axiomatic requir…

BinarizationNegation

A Reexamination of Proof Approaches for the Impossibility Theorem

2023-09-13 · Kazuya Yamamoto

The decisive-set and pivotal-voter approaches have been used to prove Arrow's impossibility theorem. This study presents a proof using a proof calculus in logic. A valid deductive inference between the premises, the axio…

valid