paper-with-me

Papers

SOTOPIA-$Ω$: Dynamic Strategy Injection Learning and Social Instruction Following Evaluation for Social Agents

2025-02-21 · Wenyuan Zhang, Tianyun Liu, Mengxiao Song, XiaoDong Li, Tingwen Liu

Despite the abundance of prior social strategies possessed by humans, there remains a paucity of research dedicated to their transfer and integration into social agents. Our proposed SOTOPIA-$\Omega$ framework aims to address and bridge this gap, with a particular focus on enhancing the social capabilities of language agents. This framework dynamically injects multi-step reasoning strategies inspired by negotiation theory and two simple direct strategies into expert agents, thereby automating the construction of a high-quality social dialogue training corpus. Additionally, we introduce the concept of Social Instruction Following (S-IF) and propose two new S-IF evaluation metrics that complement social capability. We demonstrate that several 7B models trained on high-quality corpus not only significantly surpass the expert agent (GPT-4) in achieving social goals but also enhance S-IF performance. Analysis and variant experiments validate the advantages of dynamic construction, which can especially break the agent's prolonged deadlock.

📄 PDF Abstract BibTeX arXiv:2502.15538

Code (1)

WYRipple/SOTOPIA-Omega 공식 구현

Tasks

Instruction Following

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

2023-10-18 · Xuhui Zhou, Hao Zhu, Leena Mathur, Ruohong Zhang 외

Humans are social beings; we pursue social goals in our daily interactions, which is a crucial aspect of social intelligence. Yet, AI systems' abilities in this realm remain elusive. We present SOTOPIA, an open-ended env…

ALSO: Adversarial Online Strategy Optimization for Social Agents

2026-05-15 · Xiang Li, Liping Yi, Mingze Kong, Min Zhang 외 arxiv

Social simulation provides a compelling testbed for studying social intelligence, where agents interact through multi-turn dialogues under evolving contexts and strategically adapting opponents. Such environments are inh…

Reinforcement Learning

Sotopia-RL: Reward Design for Social Intelligence

2025-08-05 · Haofei Yu, Zhengyang Qi, Yining Zhao, Kolby Nottingham 외 arxiv

Social intelligence has become a critical capability for large language models (LLMs), enabling them to engage effectively in real-world social tasks such as collaboration and negotiation. Reinforcement learning (RL) is …

Reinforcement Learning

Probabilistic Modeling of Intentions in Socially Intelligent LLM Agents

2025-10-21 · Feifan Xia, Yuyang Fang, Defang Li, Yantong Xie 외 arxiv

We present a probabilistic intent modeling framework for large language model (LLM) agents in multi-turn social dialogue. The framework maintains a belief distribution over a partner's latent intentions, initialized from…

SOTOPIA-S4: a user-friendly system for flexible, customizable, and large-scale social simulation

2025-04-19 · Xuhui Zhou, Zhe Su, Sophie Feng, Jiaxu Zhou 외

Social simulation through large language model (LLM) agents is a promising approach to explore and validate hypotheses related to social science questions and LLM agents behavior. We present SOTOPIA-S4, a fast, flexible,…

Language ModelingLanguage ModellingLarge Language ModelManagement