paper-with-me

Papers

SECURE: Semantics-aware Embodied Conversation under Unawareness for Lifelong Robot Learning

2024-09-26 · Rimvydas Rubavicius, Peter David Fagan, Alex Lascarides, Subramanian Ramamoorthy

This paper addresses a challenging interactive task learning scenario we call rearrangement under unawareness: to manipulate a rigid-body environment in a context where the agent is unaware of a concept that is key to solving the instructed task. We propose SECURE, an interactive task learning framework designed to solve such problems. It uses embodied conversation to fix its deficient domain model -- through dialogue, the agent discovers and then learns to exploit unforeseen possibilities. In particular, SECURE learns from the user's embodied corrective feedback when it makes a mistake, and it makes strategic dialogue decisions to reveal useful evidence about novel concepts for solving the instructed task. Together, these abilities allow the agent to generalise to subsequent tasks using newly acquired knowledge. We demonstrate that learning to solve rearrangement under unawareness is more data efficient when the agent is semantics-aware -- that is, during both learning and inference it augments the evidence from the user's embodied conversation with its logical consequences, stemming from semantic analysis.

📄 PDF Abstract BibTeX arXiv:2409.17755

Code (0)

등록된 구현이 없습니다.

Tasks

Novel ConceptsSentence

Similar Papers 제목 키워드 기반

Reading the Mood Behind Words: Integrating Prosody-Derived Emotional Context into Socially Responsive VR Agents

2026-03-10 · SangYeop Jeong, Yeongseo Na, Seung Gyu Jeong, Jin-Woo Jeong 외 arxiv

In VR interactions with embodied conversational agents, users' emotional intent is often conveyed more by how something is said than by what is said. However, most VR agent pipelines rely on speech-to-text processing, di…

Speech Emotion Recognition

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments

2026-04-20 · Kangan Qian, ChuChu Xie, Yang Zhong, Jingrui Pang 외 arxiv

Vision-Language-Action (VLA) models drive next-generation autonomous systems, but training them requires scalable, high-quality annotations from complex environments. Current cloud pipelines rely on generic vision-langua…

Reinforcement LearningSpatial Reasoning

DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning

2026-01-22 · Junha Lee, Eunha Park, Minsu Cho arxiv

Language-driven dexterous grasp generation requires the models to understand task semantics, 3D geometry, and complex hand-object interactions. While vision-language models have been applied to this problem, existing app…

Multimodal analysis of the predictability of hand-gesture properties

2021-08-12 · Taras Kucherenko, Rajmund Nagy, Michael Neff, Hedvig Kjellström 외

Embodied conversational agents benefit from being able to accompany their speech with gestures. Although many data-driven approaches to gesture generation have been proposed in recent years, it is still unclear whether s…

Gesture GenerationRhythm

Position: Embodied AI Requires a Privacy-Utility Trade-off

2026-05-06 · Xiaoliang Fan, Jiarui Chen, Zhuodong Liu, Ziqi Yang 외 arxiv

Embodied AI (EAI) systems are rapidly transitioning from simulations into real-world domestic and other sensitive environments. However, recent EAI solutions have largely demonstrated advancements within isolated stages …