paper-with-me

홈 › Papers

CANVAS: Commonsense-Aware Navigation System for Intuitive Human-Robot Interaction

2024-10-02 · Suhwan Choi, Yongjun Cho, Minchan Kim, JaeYoon Jung, Myunchul Joe, Yubeen Park, Minseo Kim, Sungwoong Kim, Sungjae Lee, Hwiseong Park, Jiwan Chung, Youngjae Yu

Real-life robot navigation involves more than just reaching a destination; it requires optimizing movements while addressing scenario-specific goals. An intuitive way for humans to express these goals is through abstract cues like verbal commands or rough sketches. Such human guidance may lack details or be noisy. Nonetheless, we expect robots to navigate as intended. For robots to interpret and execute these abstract instructions in line with human expectations, they must share a common understanding of basic navigation concepts with humans. To this end, we introduce CANVAS, a novel framework that combines visual and linguistic instructions for commonsense-aware navigation. Its success is driven by imitation learning, enabling the robot to learn from human navigation behavior. We present COMMAND, a comprehensive dataset with human-annotated navigation results, spanning over 48 hours and 219 km, designed to train commonsense-aware navigation systems in simulated environments. Our experiments show that CANVAS outperforms the strong rule-based system ROS NavStack across all environments, demonstrating superior performance with noisy instructions. Notably, in the orchard environment, where ROS NavStack records a 0% total success rate, CANVAS achieves a total success rate of 67%. CANVAS also closely aligns with human demonstrations and commonsense constraints, even in unseen environments. Furthermore, real-world deployment of CANVAS showcases impressive Sim2Real transfer with a total success rate of 69%, highlighting the potential of learning from human demonstrations in simulated environments for real-world applications.

📄 PDF Abstract BibTeX arXiv:2410.01273

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation LearningNavigateRobot Navigation

Similar Papers 제목 키워드 기반

MotionCanvas: Cinematic Shot Design with Controllable Image-to-Video Generation

2025-02-06 · Jinbo Xing, Long Mai, Cusuh Ham, Jiahui Huang 외

This paper presents a method that allows users to design cinematic video shots in the context of image-to-video generation. Shot design, a critical aspect of filmmaking, involves meticulously planning both camera movemen…

Image to Video GenerationVideo EditingVideo Generation

WHY: Natural Explanations from a Robot Navigator

2017-09-27 · Raj Korpan, Susan L. Epstein, Anoop Aroor, Gil Dekel

Effective collaboration between a robot and a person requires natural communication. When a robot travels with a human companion, the robot should be able to explain its navigation behavior in natural language. This pape…

Robot NavigationText Generation

EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution

2026-02-20 · Tianfu Wang, Leilei Ding, Ziyang Tao, Yi Zhan 외 arxiv

High-fidelity diagram creation requires the complex orchestration of semantic topology, visual styling, and spatial layout, posing a significant challenge for automated systems. Existing methods also suffer from a repres…

Machine Common Sense Concept Paper

2018-10-17 · David Gunning

This paper summarizes some of the technical background, research ideas, and possible development strategies for achieving machine common sense. Machine common sense has long been a critical-but-missing component of Artif…

Common Sense Reasoning

ESC: Exploration with Soft Commonsense Constraints for Zero-shot Object Navigation

2023-01-30 · Kaiwen Zhou, Kaizhi Zheng, Connor Pryor, Yilin Shen 외

The ability to accurately locate and navigate to a specific object is a crucial capability for embodied agents that operate in the real world and interact with objects to complete tasks. Such object navigation tasks usua…

Efficient ExplorationLanguage ModelingLanguage ModellingNavigate+1