paper-with-me

Papers

Representational Similarity and Model Behavior in Multi-Agent Interaction

2026-06-05 · Yujin Potter, Seun Eisape, Shiyang Lai, Alexander Huth, James Evans, Been Kim, Jacob Eisenstein, Dawn Song, Alane Suhr arxiv

Researchers have shown that neural similarity among humans predicts social closeness and cooperative success, whereas innovation often emerges from interactions among dissimilar individuals. We investigate whether these principles extend to artificial intelligence by examining interactions between large language models. In our experiments, 276 model pairs interact across eight games spanning both cooperation and novelty. We find that pairs with more similar representation spaces achieve significantly higher cooperation but exhibit reduced novelty and creativity. The effects of representational similarity on cooperation and novelty remain robust even after controlling for other factors such as performance disparity and model size. We also find that similarity in the early layers consistently shows the strongest association with cooperation and novelty, compared to the middle and later layers. This suggests that a central factor underlying these patterns could be the extent to which the two models share lexical and semantic grounding. Overall, representational similarity can be an important consideration in multi-agent system design.

📄 PDF Abstract BibTeX arXiv:2606.07818

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Hidden Coalitions in Multi-Agent AI: A Spectral Diagnostic from Internal Representations

2026-05-04 · Cameron Berg, Susan L. Schneider, Mark M. Bailey arxiv

Collections of interacting AI agents can form coalitions, creating emergent group-level organization that is critical for AI safety and alignment. However, observing agent behavior alone is often insufficient to distingu…

Multi-agent Reinforcement Learning

Hierarchical Trait-State Model for Decoding Dyadic Social Interactions

2024-11-19 · Qianying Wu, Shigeki Nakauchi, Mohammad Shehata, Shinsuke Shimojo

Traits are patterns of brain signals and behaviors that are stable over time but differ across individuals, whereas states are phasic patterns that vary over time, are influenced by the environment, yet oscillate around …

Dimensionality ReductionEEGElectroencephalogram (EEG)

Distinguishing representational geometries with controversial stimuli: Bayesian experimental design and its application to face dissimilarity judgments

2022-11-28 · Tal Golan, Wenxuan Guo, Heiko H. Schütt, Nikolaus Kriegeskorte

Comparing representations of complex stimuli in neural network layers to human brain representations or behavioral judgments can guide model development. However, even qualitatively distinct neural network models often p…

Experimental DesignFace Model

Learning Human-like Representations to Enable Learning Human Values

2023-12-21 · Andrea Wynn, Ilia Sucholutsky, Thomas L. Griffiths

How can we build AI systems that can learn any set of individual human values both quickly and safely, avoiding causing harm or violating societal standards for acceptable behavior during the learning process? We explore…

EthicsFairnessFew-Shot Learningregression+1

Lattice $φ^{4}$ field theory as a multi-agent system of financial markets

2024-11-24 · Dimitrios Bachtis

We introduce a $\phi^{4}$ lattice field theory with frustrated dynamics as a multi-agent system to reproduce stylized facts of financial markets such as fat-tailed distributions of returns and clustered volatility. Each …