paper-with-me

Papers

PHASE: PHysically-grounded Abstract Social Events for Machine Social Perception

2021-03-02 · NeurIPS Workshop SVRHM 2020 12 · Aviv Netanyahu, Tianmin Shu, Boris Katz, Andrei Barbu, Joshua B. Tenenbaum

The ability to perceive and reason about social interactions in the context of physical environments is core to human social intelligence and human-machine cooperation. However, no prior dataset or benchmark has systematically evaluated physically grounded perception of complex social interactions that go beyond short actions, such as high-fiving, or simple group activities, such as gathering. In this work, we create a dataset of physically-grounded abstract social events, PHASE, that resemble a wide range of real-life social interactions by including social concepts such as helping another agent. PHASE consists of 2D animations of pairs of agents moving in a continuous space generated procedurally using a physics engine and a hierarchical planner. Agents have a limited field of view, and can interact with multiple objects, in an environment that has multiple landmarks and obstacles. Using PHASE, we design a social recognition task and a social prediction task. PHASE is validated with human experiments demonstrating that humans perceive rich interactions in the social events, and that the simulated agents behave similarly to humans. As a baseline model, we introduce a Bayesian inverse planning approach, SIMPLE (SIMulation, Planning and Local Estimation), which outperforms state-of-the-art feed-forward neural networks. We hope that PHASE can serve as a difficult new challenge for developing new models that can recognize complex social interactions.

📄 PDF Abstract BibTeX arXiv:2103.01933

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Grounding Social Perception in Intuitive Physics

2026-03-28 · Lance Ying, Aydan Y. Huang, Aviv Netanyahu, Andrei Barbu 외 arxiv

People infer rich social information from others' actions. These inferences are often constrained by the physical world: what agents can do, what obstacles permit, and how the physical actions of agents causally change a…

Physically Native World Models: A Hamiltonian Perspective on Generative World Modeling

2026-05-01 · Sen Cui, Jingheng Ma arxiv

World models have recently re-emerged as a central paradigm for embodied intelligence, robotics, autonomous driving, and model-based reinforcement learning. However, current world model research is often dominated by thr…

Reinforcement LearningAutonomous DrivingDecision Making

Can World Simulators Reason? Gen-ViRe: A Generative Visual Reasoning Benchmark

2025-11-17 · Xinxin Liu, Zhaopan Xu, Ming Li, Kai Wang 외 arxiv

While Chain-of-Thought (CoT) prompting enables sophisticated symbolic reasoning in LLMs, it remains confined to discrete text and cannot simulate the continuous, physics-governed dynamics of the real world. Recent video …

Visual ReasoningVideo Generation

Social Structure Matters in 3D Human-Human Interaction Generation

2026-06-23 · Zhongju Wang, Beier Wang, Yatao Bian, Pichao Wang 외 arxiv

Although text-to-motion generation has achieved strong progress in synthesizing realistic single-person motions from language, extending it to text-driven 3D human-human interaction (HHI) remains non-trivial, as HHI requ…

MeGAS: Thermomechanical Dynamic Gaussian Splatting for Thermophysical Scene Editing

2026-06-22 · Zesong Yang, Yuanhang Lei, Liyuan Cui, Yihang Chen 외 arxiv

Recent advances integrate physically grounded Newtonian dynamics with neural rendering frameworks, narrowing the gap between photorealistic scene reconstruction and physics-based animation. However, existing approaches f…