paper-with-me

홈 › Papers

Sim-2-Sim Transfer for Vision-and-Language Navigation in Continuous Environments

2022-04-20 · Jacob Krantz, Stefan Lee

Recent work in Vision-and-Language Navigation (VLN) has presented two environmental paradigms with differing realism -- the standard VLN setting built on topological environments where navigation is abstracted away, and the VLN-CE setting where agents must navigate continuous 3D environments using low-level actions. Despite sharing the high-level task and even the underlying instruction-path data, performance on VLN-CE lags behind VLN significantly. In this work, we explore this gap by transferring an agent from the abstract environment of VLN to the continuous environment of VLN-CE. We find that this sim-2-sim transfer is highly effective, improving over the prior state of the art in VLN-CE by +12% success rate. While this demonstrates the potential for this direction, the transfer does not fully retain the original performance of the agent in the abstract setting. We present a sequence of experiments to identify what differences result in performance degradation, providing clear directions for further improvement.

📄 PDF Abstract BibTeX arXiv:2204.09667

Code (0)

등록된 구현이 없습니다.

Tasks

NavigateVision and Language Navigation

Similar Papers 제목 키워드 기반

Beyond the Nav-Graph: Vision-and-Language Navigation in Continuous Environments

2020-04-06 · ECCV 2020 8 · Jacob Krantz, Erik Wijmans, Arjun Majumdar, Dhruv Batra 외

We develop a language-guided navigation task set in a continuous 3D environment where agents must execute low-level actions to follow natural language navigation directions. By being situated in continuous environments, …

Vision and Language Navigation

Bridging the Gap Between Learning in Discrete and Continuous Environments for Vision-and-Language Navigation

2022-03-05 · CVPR 2022 1 · Yicong Hong, Zun Wang, Qi Wu, Stephen Gould

Most existing works in vision-and-language navigation (VLN) focus on either discrete or continuous environments, training agents that cannot generalize across the two. The fundamental difference between the two setups is…

Imitation LearningVision and Language Navigation

Sim-to-Real Transfer for Vision-and-Language Navigation

2020-11-07 · Peter Anderson, Ayush Shrivastava, Joanne Truong, Arjun Majumdar 외

We study the challenging problem of releasing a robot in a previously unseen environment, and having it follow unconstrained natural language navigation instructions. Recent work on the task of Vision-and-Language Naviga…

Vision and Language Navigation

ETPNav: Evolving Topological Planning for Vision-Language Navigation in Continuous Environments

2023-04-06 · Dong An, Hanqing Wang, Wenguan Wang, Zun Wang 외

Vision-language navigation is a task that requires an agent to follow instructions to navigate in environments. It becomes increasingly crucial in the field of embodied AI, with potential applications in autonomous navig…

Autonomous NavigationNavigateVision-Language Navigation

Beyond the Nav-Graph: Vision-and-Language Navigation in Continuous Environments – Extended Abstract

2020-06-12 · ICML Workshop LaReL 2020 7 · Jacob Krantz, Erik Wijmans, Arjun Majumdar, Dhruv Batra 외

We develop a language-guided navigation task set in a continuous 3D environment where agents must execute low-level actions to follow natural language navigation directions. By being situated in continuous environments, …

Vision and Language Navigation