paper-with-me

Papers

FreeAskWorld: An Interactive and Closed-Loop Simulator for Human-Centric Embodied AI

2025-11-17 · Yuhang Peng, Yizhou Pan, Xinning He, Jihaoyu Yang, Xinyu Yin, Han Wang, Xiaoji Zheng, Chao Gao, Jiangtao Gong arxiv

As embodied intelligence emerges as a core frontier in artificial intelligence research, simulation platforms must evolve beyond low-level physical interactions to capture complex, human-centered social behaviors. We introduce FreeAskWorld, an interactive simulation framework that integrates large language models (LLMs) for high-level behavior planning and semantically grounded interaction, informed by theories of intention and social cognition. Our framework supports scalable, realistic human-agent simulations and includes a modular data generation pipeline tailored for diverse embodied tasks. To validate the framework, we extend the classic Vision-and-Language Navigation (VLN) task into a interaction enriched Direction Inquiry setting, wherein agents can actively seek and interpret navigational guidance. We present and publicly release FreeAskWorld, a large-scale benchmark dataset comprising reconstructed environments, six diverse task types, 16 core object categories, 63,429 annotated sample frames, and more than 17 hours of interaction data to support training and evaluation of embodied AI systems. We benchmark VLN models, and human participants under both open-loop and closed-loop settings. Experimental results demonstrate that models fine-tuned on FreeAskWorld outperform their original counterparts, achieving enhanced semantic understanding and interaction competency. These findings underscore the efficacy of socially grounded simulation frameworks in advancing embodied AI systems toward sophisticated high-level planning and more naturalistic human-agent interaction. Importantly, our work underscores that interaction itself serves as an additional information modality.

📄 PDF Abstract BibTeX arXiv:2511.13524

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On Learning Closed-Loop Probabilistic Multi-Agent Simulator

2025-08-01 · Juanwu Lu, Rohit Gupta, Ahmadreza Moradipari, Kyungtae Han 외 arxiv

The rapid iteration of autonomous vehicle (AV) deployments leads to increasing needs for building realistic and scalable multi-agent traffic simulators for efficient evaluation. Recent advances in this area focus on clos…

Trajectory PredictionBayesian Inference

RIFT: Closed-Loop RL Fine-Tuning for Realistic and Controllable Traffic Simulation

2025-05-06 · Keyu Chen, Wenchao Sun, Hao Cheng, Sifa Zheng

Achieving both realism and controllability in interactive closed-loop traffic simulation remains a key challenge in autonomous driving. Data-driven simulation methods reproduce realistic trajectories but suffer from cova…

Autonomous DrivingImitation Learning

nuCarla: A nuScenes-Style Bird's-Eye View Perception Dataset for CARLA Simulation

2025-11-12 · Zhijie Qiao, Zhong Cao, Henry X. Liu arxiv

End-to-end (E2E) autonomous driving heavily relies on closed-loop simulation, where perception, planning, and control are jointly trained and evaluated in interactive environments. Yet, most existing datasets are collect…

Autonomous Driving

CausalDrive: Real-time Causal World Models for Autonomous Driving

2026-06-13 · Tianyi Yan, Huan Zheng, Dubing Chen, Meizhi Qu 외 arxiv

World models have emerged as a promising paradigm for scaling autonomous driving (AD) data, yet existing video generative models fall short as interactive simulators. Layout-conditioned renderers rely on "oracle" future …

Reinforcement LearningAutonomous Driving

Unraveling the Effects of Synthetic Data on End-to-End Autonomous Driving

2025-03-23 · Junhao Ge, Zuhong Liu, Longteng Fan, Yifan Jiang 외

End-to-end (E2E) autonomous driving (AD) models require diverse, high-quality data to perform well across various driving scenarios. However, collecting large-scale real-world data is expensive and time-consuming, making…

3DGSAutonomous DrivingDiversityNeRF+1