paper-with-me

Papers

Act on What You See: Unlocking Safe Social Navigation in Vision-Language-Action Models

2026-06-09 · Qingzi Wang, Xiyang Wu, Guangyao Shi, Dianwei Chen, Xianfeng Yang, Dinesh Manocha arxiv

Safe social navigation requires robots to distinguish people from ordinary obstacles and to react before danger becomes imminent. We show that pretrained Vision-Language-Action (VLA) models already encode pedestrian-object distinctions and future collision signals in their internal representations, but behavior cloning fails to translate these signals into socially appropriate actions. To address this mismatch, we propose SALSA, a two-stage annotation-free post-training framework: (1) social behavioral alignment bridges intermediate-layer social features to the action head and trains on counterfactual human-object scene pairs to break visual saliency shortcuts; (2) temporal safety alignment provides automatically generated future-risk supervision to enable anticipatory collision avoidance. On SCAND and real-world deployment, SALSA reduces near-collisions by 86.4% and improves social counterfactual accuracy from 53% to 93%, demonstrating that safer social navigation can be achieved by teaching VLA policies to act on representations they already possess. These results show that pretrained VLA policies can be adapted for safer social navigation by better aligning their latent representations with action generation.

📄 PDF Abstract BibTeX arXiv:2606.10495

Code (0)

등록된 구현이 없습니다.

Tasks

Collision Avoidance

Similar Papers 제목 키워드 기반

Socially Aware Motion Planning with Deep Reinforcement Learning

2017-03-26 · Yu Fan Chen, Michael Everett, Miao Liu, Jonathan P. How

For robotic vehicles to navigate safely and efficiently in pedestrian-rich environments, it is important to model subtle human behaviors and navigation rules (e.g., passing on the right). However, while instinctive to hu…

Autonomous NavigationDeep Reinforcement LearningMotion PlanningNavigate+3

SONG: A Photorealistic 3D Gaussian Simulation Platform for Benchmarking Social Navigation

2026-07-28 · Weiqi Huang, Dianyi Yang, Jiaxin Li, Shuangyi Dong 외 arxiv

Social navigation has progressed from simplified 2D environments toward a more general vision-based setting, in which a robot needs to achieve socially compliant behavior purely from onboard visual observations. Yet supp…

SocialNav-MoE: A Mixture-of-Experts Vision Language Model for Socially Compliant Navigation with Reinforcement Fine-Tuning

2025-12-15 · Tomohito Kawabata, Xinyu Zhang, Ling Xiao arxiv

For robots navigating in human-populated environments, safety and social compliance are equally critical, yet prior work has mostly emphasized safety. Socially compliant navigation that accounts for human comfort, social…

Semantic Similarity

Walk With Me: Long-Horizon Social Navigation for Human-Centric Outdoor Assistance

2026-04-29 · Lingfeng Zhang, Xiaoshuai Hao, Xizhou Bu, Yingbo Tang 외 arxiv

Assisting humans in open-world outdoor environments requires robots to translate high-level natural-language intentions into safe, long-horizon, and socially compliant navigation behavior. Existing map-based methods rely…

MAction-SocialNav: Multi-Action Socially Compliant Navigation via Reasoning-enhanced Prompt Tuning

2025-12-25 · Zishuo Wang, Xinyu Zhang, Zhuonan Liu, Tomohito Kawabata 외 arxiv

Socially compliant navigation requires robots to move safely and appropriately in human-centered environments by respecting social norms. However, social norms are often ambiguous, and in a single scenario, multiple acti…

Robot Navigation