paper-with-me

홈 › Papers

S$^3$IT: A Benchmark for Spatially Situated Social Intelligence Test

2025-12-23 · Zhe Sun, Xueyuan Yang, Yujie Lu, Zhenliang Zhang arxiv

The integration of embodied agents into human environments demands embodied social intelligence: reasoning over both social norms and physical constraints. However, existing evaluations fail to address this integration, as they are limited to either disembodied social reasoning (e.g., in text) or socially-agnostic physical tasks. Both approaches fail to assess an agent's ability to integrate and trade off both physical and social constraints within a realistic, embodied context. To address this challenge, we introduce Spatially Situated Social Intelligence Test (S$^{3}$IT), a benchmark specifically designed to evaluate embodied social intelligence. It is centered on a novel and challenging seat-ordering task, requiring an agent to arrange seating in a 3D environment for a group of large language model-driven (LLM-driven) NPCs with diverse identities, preferences, and intricate interpersonal relationships. Our procedurally extensible framework generates a vast and diverse scenario space with controllable difficulty, compelling the agent to acquire preferences through active dialogue, perceive the environment via autonomous exploration, and perform multi-objective optimization within a complex constraint network. We evaluate state-of-the-art LLMs on S$^{3}$IT and found that they still struggle with this problem, showing an obvious gap compared with the human baseline. Results imply that LLMs have deficiencies in spatial intelligence, yet simultaneously demonstrate their ability to achieve near human-level competence in resolving conflicts that possess explicit textual cues.

📄 PDF Abstract BibTeX arXiv:2512.19992

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Embodied, Situated, and Grounded Intelligence: Implications for AI

2022-10-24 · Tyler Millhouse, Melanie Moses, Melanie Mitchell

In April of 2022, the Santa Fe Institute hosted a workshop on embodied, situated, and grounded intelligence as part of the Institute's Foundations of Intelligence project. The workshop brought together computer scientist…

ToM-SSI: Evaluating Theory of Mind in Situated Social Interactions

2025-09-05 · Matteo Bortoletto, Constantin Ruhdorfer, Andreas Bulling arxiv

Most existing Theory of Mind (ToM) benchmarks for foundation models rely on variations of the Sally-Anne test, offering only a very limited perspective on ToM and neglecting the complexity of human social interactions. T…

Accelerating the Development of Multimodal, Integrative-AI Systems with Platform for Situated Intelligence

2020-10-12 · Sean Andrist, Dan Bohus

We describe Platform for Situated Intelligence, an open-source framework for multimodal, integrative-AI systems. The framework provides infrastructure, tools, and components that enable and accelerate the development of …

Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?

2025-12-19 · Zabir Al Nazi, GM Shahariar, Md. Abrar Hossain, Wei Peng arxiv

Theory of Mind (ToM) - the ability to attribute beliefs and intents to others - is fundamental for social intelligence, yet Vision-Language Model (VLM) evaluations remain largely Western-centric. In this work, we introdu…

Where Norms and References Collide: Evaluating LLMs on Normative Reasoning

2026-02-03 · Mitchell Abrams, Kaveh Eskandari Miandoab, Felix Gervits, Vasanth Sarathy 외 arxiv

Embodied agents, such as robots, will need to interact in situated environments where successful communication often depends on reasoning over social norms: shared expectations that constrain what actions are appropriate…