paper-with-me

홈 › Papers

Social-R1: Towards Human-like Social Reasoning in LLMs

2026-03-10 · Jincenzi Wu, Yuxuan Lei, Jianxun Lian, Yitian Huang, Lexin Zhou, Haotian Li, Xing Xie, Helen Meng arxiv

While large language models demonstrate remarkable capabilities across numerous domains, social intelligence - the capacity to perceive social cues, infer mental states, and generate appropriate responses - remains a critical challenge, particularly for enabling effective human-AI collaboration and developing AI that truly serves human needs. Current models often rely on superficial patterns rather than genuine social reasoning. We argue that cultivating human-like social intelligence requires training with challenging cases that resist shortcut solutions. To this end, we introduce ToMBench-Hard, an adversarial benchmark designed to provide hard training examples for social reasoning. Building on this, we propose Social-R1, a reinforcement learning framework that aligns model reasoning with human cognition through multi-dimensional rewards. Unlike outcome-based RL, Social-R1 supervises the entire reasoning process, enforcing structural alignment, logical integrity, and information density. Results show that our approach enables a 4B parameter model to surpass much larger counterparts and generalize robustly across eight diverse benchmarks. These findings demonstrate that challenging training cases with trajectory-level alignment offer a path toward efficient and reliable social intelligence.

📄 PDF Abstract BibTeX arXiv:2603.09249

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Can MLLMs Read the Room? A Multimodal Benchmark for Assessing Deception in Multi-Party Social Interactions

2025-11-20 · Caixin Kang, Yifei Huang, Liangyang Ouyang, Mingfang Zhang 외 arxiv

Despite their advanced reasoning capabilities, state-of-the-art Multimodal Large Language Models (MLLMs) demonstrably lack a core component of human intelligence: the ability to `read the room' and assess deception in co…

SocialGaze: Improving the Integration of Human Social Norms in Large Language Models

2024-10-11 · Anvesh Rao Vijjini, Rakesh R. Menon, Jiayi Fu, Shashank Srivastava 외

While much research has explored enhancing the reasoning capabilities of large language models (LLMs) in the last few years, there is a gap in understanding the alignment of these models with social values and norms. We …

Language ModelingLanguage Modelling

Effects of Theory of Mind and Prosocial Beliefs on Steering Human-Aligned Behaviors of LLMs in Ultimatum Games

2025-05-30 · Neemesh Yadav, Palakorn Achananuparp, Jing Jiang, Ee-Peng Lim

Large Language Models (LLMs) have shown potential in simulating human behaviors and performing theory-of-mind (ToM) reasoning, a crucial skill for complex social interactions. In this study, we investigate the role of To…

Decision Making

Social-LLaVA: Enhancing Robot Navigation through Human-Language Reasoning in Social Spaces

2024-12-30 · Amirreza Payandeh, Daeun Song, Mohammad Nazeri, Jing Liang 외

Most existing social robot navigation techniques either leverage hand-crafted rules or human demonstrations to connect robot perception to socially compliant actions. However, there remains a significant gap in effective…

2kRobot NavigationVisual Question Answering (VQA)

Logic-Guided Socially-aware Robot Navigation World Model

2025-10-27 · Weizheng Wang, Obi Ike, Soyun Choi, Sungeun Hong 외 arxiv

Social robot navigation increasingly relies on large language models for reasoning, path planning, and enabling movement in dynamic human spaces. However, relying solely on LLMs for planning often leads to unpredictable …

Collision AvoidanceRobot Navigation