paper-with-me

Papers

Investigating social alignment via mirroring in a system of interacting language models

2024-12-07 · Harvey McGuinness, Tianyu Wang, Carey E. Priebe, Hayden Helm

Alignment is a social phenomenon wherein individuals share a common goal or perspective. Mirroring, or mimicking the behaviors and opinions of another individual, is one mechanism by which individuals can become aligned. Large scale investigations of the effect of mirroring on alignment have been limited due to the scalability of traditional experimental designs in sociology. In this paper, we introduce a simple computational framework that enables studying the effect of mirroring behavior on alignment in multi-agent systems. We simulate systems of interacting large language models in this framework and characterize overall system behavior and alignment with quantitative measures of agent dynamics. We find that system behavior is strongly influenced by the range of communication of each agent and that these effects are exacerbated by increased rates of mirroring. We discuss the observed simulated system behavior in the context of known human social dynamics.

📄 PDF Abstract BibTeX arXiv:2412.06834

Code (0)

등록된 구현이 없습니다.

Tasks

Sociology

Similar Papers 제목 키워드 기반

Modelling Human Values for AI Reasoning

2024-02-09 · Nardine Osman, Mark d'Inverno

One of today's most significant societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacting agents (human and artificial), aligns with human values. To addres…

Neural Synchrony Between Socially Interacting Language Models

2026-02-19 · Zhining Zhang, Wentao Zhu, Chi Han, Yizhou Wang 외 arxiv

Neuroscience has uncovered a fundamental mechanism of our social nature: human brain activity becomes synchronized with others in many social contexts involving interaction. Traditionally, social minds have been regarded…

TriAlign: Towards Universal Truth Consistency in Personalized LLM Alignment

2026-06-01 · Thi-Nhung Nguyen, Linhao Luo, Rollin Omari, Junae Kim 외 arxiv

Personalized large language models adapt responses to users' preferences and social attributes, but can introduce substantial universal truth inconsistencies across social groups, where some groups systematically receive…

Multi-agent Reinforcement Learning

Towards Social Autonomous Vehicles: Efficient Collision Avoidance Scheme Using Richardson's Arms Race Model

2017-08-06 · Faisal Riaz, Muaz A. Niazi

Background Road collisions and casualties pose a serious threat to commuters around the globe. Autonomous Vehicles (AVs) aim to make the use of technology to reduce the road accidents. However, the most of research work …

Autonomous VehiclesCollision Avoidance

Psychological and behavioural responses in human-agent vs. human-human interactions: a systematic review and meta-analysis

2025-09-25 · Jianan Zhou, Fleur Corbett, Joori Byun, Talya Porat 외 arxiv

Interactive intelligent agents are being integrated across society. Despite achieving human-like capabilities, humans' responses to these agents remain poorly understood, with research fragmented across disciplines. We c…