paper-with-me

홈 › Papers

Hidden Coalitions in Multi-Agent AI: A Spectral Diagnostic from Internal Representations

2026-05-04 · Cameron Berg, Susan L. Schneider, Mark M. Bailey arxiv

Collections of interacting AI agents can form coalitions, creating emergent group-level organization that is critical for AI safety and alignment. However, observing agent behavior alone is often insufficient to distinguish genuine informational coupling from spurious similarity, as consequential coalitions may form at the level of internal representations before any overt behavioral change is apparent. Here, we introduce a practical method for detecting coalition structure from the internal neural representations of multi-agent systems. The approach constructs a pairwise mutual-information graph from the hidden states of agents and applies spectral partitioning to identify the most salient coalition boundary. We validate this method in two domains. First, in multi-agent reinforcement learning environments, the method successfully recovers programmed hierarchical and dynamic coalition structures and correctly rejects false positives arising from behavioral coordination without informational coupling. Second, using a large language model, the method identifies coalition structures implied by descriptive prompts, tracks dynamic team reassignments, and reveals a representational hierarchy where explicit labels dominate over conflicting interaction patterns. Across both settings, the recovered partition reveals subgroup organization that a scalar cross-agent mutual-information measure cannot distinguish. The results demonstrate that analyzing hidden-state mutual information through spectral partitioning provides a scalable diagnostic for identifying representational coalitions, offering a valuable tool for monitoring emergent structure in distributed AI systems.

📄 PDF Abstract BibTeX arXiv:2605.06696

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

Nucleolus Credit Assignment for Effective Coalitions in Multi-agent Reinforcement Learning

2025-03-01 · Yugu Li, Zehong Cao, Jianglin Qiao, Siyi Hu

In cooperative multi-agent reinforcement learning (MARL), agents typically form a single grand coalition based on credit assignment to tackle a composite task, often resulting in suboptimal performance. This paper propos…

Multi-agent Reinforcement LearningQ-LearningStarcraft

Visualizing Coalition Formation: From Hedonic Games to Image Segmentation

2026-03-09 · Pedro Henrique de Paula França, Lucas Lopes Felipe, Daniel Sadoc Menasché arxiv

We propose image segmentation as a visual diagnostic testbed for coalition formation in hedonic games. Modeling pixels as agents on a graph, we study how a granularization parameter shapes equilibrium fragmentation and b…

Image Segmentation

Contextual and Possibilistic Reasoning for Coalition Formation

2020-06-19 · Antonis Bikakis, Patrice Caire

In multiagent systems, agents often have to rely on other agents to reach their goals, for example when they lack a needed resource or do not have the capability to perform a required action. Agents therefore need to coo…

Offline Learning of Nash Stable Coalition Structures with Possibly Overlapping Coalitions

2026-02-15 · Saar Cohen arxiv

Coalition formation concerns strategic collaborations of selfish agents that form coalitions based on their preferences. It is often assumed that coalitions are disjoint and preferences are fully known, which may not hol…

Solidarity to achieve stability

2023-02-15 · Jorge Alcalde-Unzu, Oihane Gallo, Elena Inarra, Juan D. Moreno-Ternero

Agents may form coalitions. Each coalition shares its endowment among its agents by applying a sharing rule. The sharing rule induces a coalition formation problem by assuming that agents rank coalitions according to the…