paper-with-me

홈 › Papers

Versatile Navigation under Partial Observability via Value-guided Diffusion Policy

2024-04-01 · CVPR 2024 1 · Gengyu Zhang, Hao Tang, Yan Yan

Route planning for navigation under partial observability plays a crucial role in modern robotics and autonomous driving. Existing route planning approaches can be categorized into two main classes: traditional autoregressive and diffusion-based methods. The former often fails due to its myopic nature, while the latter either assumes full observability or struggles to adapt to unfamiliar scenarios, due to strong couplings with behavior cloning from experts. To address these deficiencies, we propose a versatile diffusion-based approach for both 2D and 3D route planning under partial observability. Specifically, our value-guided diffusion policy first generates plans to predict actions across various timesteps, providing ample foresight to the planning. It then employs a differentiable planner with state estimations to derive a value function, directing the agent's exploration and goal-seeking behaviors without seeking experts while explicitly addressing partial observability. During inference, our policy is further enhanced by a best-plan-selection strategy, substantially boosting the planning success rate. Moreover, we propose projecting point clouds, derived from RGB-D inputs, onto 2D grid-based bird-eye-view maps via semantic segmentation, generalizing to 3D environments. This simple yet effective adaption enables zero-shot transfer from 2D-trained policy to 3D, cutting across the laborious training for 3D policy, and thus certifying our versatility. Experimental results demonstrate our superior performance, particularly in navigating situations beyond expert demonstrations, surpassing state-of-the-art autoregressive and diffusion-based baselines for both 2D and 3D scenarios.

📄 PDF Abstract BibTeX arXiv:2404.02176

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingSemantic Segmentation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Deep Visual Navigation under Partial Observability

2021-09-16 · Bo Ai, Wei Gao, Vinay, David Hsu

How can a robot navigate successfully in rich and diverse environments, indoors or outdoors, along office corridors or trails on the grassland, on the flat ground or the staircase? To this end, this work aims to address …

Imitation LearningNavigateVisual Navigation

LLMs for Text-Based Exploration and Navigation Under Partial Observability

2026-03-10 · Stephan Sandfuchs, Maximilian Melchert, Jörg Frochte arxiv

Exploration and goal-directed navigation in unknown layouts are central to inspection, logistics, and search-and-rescue. We ask whether large language models (LLMs) can function as \emph{text-only} controllers under part…

Program Synthesis

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability

2026-05-14 · Yushen Liu, Yin-Jen Chen, Ziyi Chen, Tao Wang 외 arxiv

Many safety-critical control problems are modeled as risk-sensitive partially observable Markov decision processes, where the controller must make decisions from incomplete observations while balancing task performance a…

Reinforcement Learning

Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability

2026-05-21 · Taewoon Kim, Vincent François-Lavet, Michael Cochez arxiv

Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly model short-term-to-long-term transfer of symbolic observations. We stud…

Reinforcement LearningKnowledge Graphs

Integrating Deep RL and Bayesian Inference for ObjectNav in Mobile Robotics

2026-03-26 · João Castelo-Branco, José Santos-Victor, Alexandre Bernardino arxiv

Autonomous object search is challenging for mobile robots operating in indoor environments due to partial observability, perceptual uncertainty, and the need to trade off exploration and navigation efficiency. Classical …

Reinforcement LearningBayesian Inference