paper-with-me

Papers

Causal Foundations of Collective Agency

2026-04-30 · Frederik Hytting Jørgensen, Sebastian Weichwald, Lewis Hammond arxiv

A key challenge for the safety of advanced AI systems is the possibility that multiple simpler agents might inadvertently form a collective agent with capabilities and goals distinct from those of any individual. More generally, determining when a group of agents can be viewed as a unified collective agent is a foundational question in the study of interactions and incentives in both biological and artificial systems. We adopt a behavioral perspective in answering this question, ascribing collective agency to a group when viewing the group's joint actions as rational and goal-directed successfully predicts its behavior. We formalize this perspective on collective agency using causal games -- which are causal models of strategic, multi-agent interactions -- and causal abstraction -- which formalizes when a simple, high-level model faithfully captures a more complex, low-level model. We use this framework to solve a puzzle regarding multi-agent incentives in actor-critic models and to make quantitative assessments of the degree of collective agency exhibited by different voting mechanisms. Our framework aims to provide a foundation for theoretical and empirical work to understand, predict, and control emergent collective agents in multi-agent AI systems.

📄 PDF Abstract BibTeX arXiv:2605.00248

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Pareto Optimality, Functional Dependence and Collective Agency

2021-04-19 · Chenwei Shi, Yiyang Wang

This paper approaches the problem of understanding collective agency from a logical and game-theoretical perspective. Instead of collective intentionality, our analysis highlights the role of Pareto optimality. To facili…

Intent-aligned AI systems deplete human agency: the need for agency foundations research in AI safety

2023-05-30 · Catalin Mitelut, Ben Smith, Peter Vamplew

The rapid advancement of artificial intelligence (AI) systems suggests that artificial general intelligence (AGI) systems may soon arrive. Many researchers are concerned that AIs and AGIs will harm humans via intentional…

A three-dimensional typology of agency for advanced AI systems

2026-08-20 · Willem Fourie arxiv

Research on the agency of advanced artificial intelligence (AI) systems focuses on agency as a normative concept and on the agency of particularly agentic AI systems. While recent work also focuses on the different profi…

Accountable Human-AI Deliberation with LLMs: Scaling Collective Intelligence through Symbiotic Scaffolding

2026-05-26 · Wajdi Zaghouani arxiv

Large language models (LLMs) can support democratic deliberation at scales previously constrained by turn-taking and facilitation bandwidth. Recent work shows that LLM-generated group statements are often preferred over …

Adversarial Robustness

Modelling collective motion based on the principle of agency

2017-12-04 · Katja Ried, Thomas Müller, Hans J. Briegel

Collective motion is an intriguing phenomenon, especially considering that it arises from a set of simple rules governing local interactions between individuals. In theoretical models, these rules are normally \emph{assu…