paper-with-me

Papers

Spatial Concept-Based Navigation with Human Speech Instructions via Probabilistic Inference on Bayesian Generative Model

2020-02-18 · Akira Taniguchi, Yoshinobu Hagiwara, Tadahiro Taniguchi, Tetsunari Inamura

Robots are required to not only learn spatial concepts autonomously but also utilize such knowledge for various tasks in a domestic environment. Spatial concept represents a multimodal place category acquired from the robot's spatial experience including vision, speech-language, and self-position. The aim of this study is to enable a mobile robot to perform navigational tasks with human speech instructions, such as `Go to the kitchen', via probabilistic inference on a Bayesian generative model using spatial concepts. Specifically, path planning was formalized as the maximization of probabilistic distribution on the path-trajectory under speech instruction, based on a control-as-inference framework. Furthermore, we described the relationship between probabilistic inference based on the Bayesian generative model and control problem including reinforcement learning. We demonstrated path planning based on human instruction using acquired spatial concepts to verify the usefulness of the proposed approach in the simulator and in real environments. Experimentally, places instructed by the user's speech commands showed high probability values, and the trajectory toward the target place was correctly estimated. Our approach, based on probabilistic inference concerning decision-making, can lead to further improvement in robot autonomy.

📄 PDF Abstract BibTeX arXiv:2002.07381

Code (1)

a-taniguchi/SpCoNavi 공식 구현

Tasks

Decision MakingReinforcement Learning

Similar Papers 제목 키워드 기반

Hierarchical Path-planning from Speech Instructions with Spatial Concept-based Topometric Semantic Mapping

2022-03-21 · Akira Taniguchi, Shuya Ito, Tadahiro Taniguchi

Assisting individuals in their daily activities through autonomous mobile robots, especially for users without specialized knowledge, is crucial. Specifically, the capability of robots to navigate to destinations based o…

Navigate

Talk2Nav: Long-Range Vision-and-Language Navigation with Dual Attention and Spatial Memory

2019-10-04 · Arun Balajee Vasudevan, Dengxin Dai, Luc van Gool

The role of robots in society keeps expanding, bringing with it the necessity of interacting and communicating with humans. In order to keep such interaction intuitive, we provide automatic wayfinding based on verbal nav…

Autonomous DrivingVision and Language NavigationVisual Navigation

NAVCON: A Cognitively Inspired and Linguistically Grounded Corpus for Vision and Language Navigation

2024-12-17 · Karan Wanchoo, Xiaoye Zuo, Hannah Gonzalez, Soham Dan 외

We present NAVCON, a large-scale annotated Vision-Language Navigation (VLN) corpus built on top of two popular datasets (R2R and RxR). The paper introduces four core, cognitively motivated and linguistically grounded, na…

Few-Shot LearningVision and Language NavigationVision-Language Navigation

The OFAI Multi-Modal Task Description Corpus

2016-05-01 · LREC 2016 5 · Stephanie Schreitter, Brigitte Krenn

The OFAI Multimodal Task Description Corpus (OFAI-MMTD Corpus) is a collection of dyadic teacher-learner (human-human and human-robot) interactions. The corpus is multimodal and tracks the communication signals exchanged…

$NavA^3$: Understanding Any Instruction, Navigating Anywhere, Finding Anything

2025-08-06 · Lingfeng Zhang, Xiaoshuai Hao, Yingbo Tang, Haoxiang Fu 외 arxiv

Embodied navigation is a fundamental capability of embodied intelligence, enabling robots to move and interact within physical environments. However, existing navigation tasks primarily focus on predefined object navigat…

Instruction FollowingObject Localization