paper-with-me

홈 › Papers

On the Strengths and Weaknesses of Data for Open-set Embodied Assistance

2026-03-05 · Pradyumna Tambwekar, Andrew Silva, Deepak Gopinath, Jonathan DeCastro, Xiongyi Cui, Guy Rosman arxiv

Embodied foundation models are increasingly performant in real-world domains such as robotics or autonomous driving. These models are often deployed in interactive or assistive settings, where it is important that these assistive models generalize to new users and new tasks. Diverse interactive data generation offers a promising avenue for providing data-efficient generalization capabilities for interactive embodied foundation models. In this paper, we investigate the generalization capabilities of a multimodal foundation model fine-tuned on diverse interactive assistance data in a synthetic domain. We explore generalization along two axes: a) assistance with unseen categories of user behavior and b) providing guidance in new configurations not encountered during training. We study a broad capability called \textbf{Open-Set Corrective Assistance}, in which the model needs to inspect lengthy user behavior and provide assistance through either corrective actions or language-based feedback. This task remains unsolved in prior work, which typically assumes closed corrective categories or relies on external planners, making it a challenging testbed for evaluating the limits of assistive data. To support this task, we generate synthetic assistive datasets in Overcooked and fine-tune a LLaMA-based model to evaluate generalization to novel tasks and user behaviors. Our approach provides key insights into the nature of assistive datasets required to enable open-set assistive intelligence. In particular, we show that performant models benefit from datasets that cover different aspects of assistance, including multimodal grounding, defect inference, and exposure to diverse scenarios.

📄 PDF Abstract BibTeX arXiv:2603.04819

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Similar Papers 제목 키워드 기반

Open6DOR: Benchmarking Open-instruction 6-DoF Object Rearrangement and A VLM-based Approach

2024-10-24 · IROS2024 2024 10 · Yufei Ding, Haoran Geng, Chaoyi Xu, Xiaomeng Fang 외

In this work, we propel the pioneer construction of the benchmark and approach for table-top Open-instruction 6-DoF Object Rearrangement (Open6DOR). Specifically, we collect a synthetic dataset of 200+ objects and carefu…

BenchmarkingInstruction FollowingObject Rearrangement

Quantifying Morphological Computation

2013-01-29 · Keyan Zahedi, Nihat Ay

The field of embodied intelligence emphasises the importance of the morphology and environment with respect to the behaviour of a cognitive system. The contribution of the morphology to the behaviour, commonly known as m…

Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making

2024-10-09 · Manling Li, Shiyu Zhao, Qineng Wang, Kangrui Wang 외

We aim to evaluate Large Language Models (LLMs) for embodied decision making. While a significant body of work has been leveraging LLMs for decision making in embodied environments, we still lack a systematic understandi…

BenchmarkingDecision MakingHallucination

Cricket Player Profiling: Unraveling Strengths and Weaknesses Using Text Commentary Data

2023-11-12 · Swarup Ranjan Behera, Vijaya V. Saradhi

Devising player-specific strategies in cricket necessitates a meticulous understanding of each player's unique strengths and weaknesses. Nevertheless, the absence of a definitive computational approach to extract such in…

Dimensionality Reduction

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions

2024-12-07 · Ola Shorinwa, Zhiting Mei, Justin Lidard, Allen Z. Ren 외

The remarkable performance of large language models (LLMs) in content generation, coding, and common-sense reasoning has spurred widespread integration into many facets of society. However, integration of LLMs raises val…

ChatbotCommon Sense ReasoningUncertainty Quantificationvalid