paper-with-me

Papers

MVP-Nav: Multi-layer Value Map Planner Navigator

2026-06-30 · Wenyuan Xie, Shaokai Wu, Yijin Zhou, Yanbiao Ji, Guodong Zhang, Bayram Bayramli, Qiuchang Li, Xunchu Zhou, Yue Ding, Hongtao Lu arxiv

Zero-shot Object Goal Navigation (ZSON) with RGB-only perception poses a fundamental challenge for embodied agents, as the absence of explicit depth information introduces severe physical uncertainty and semantic-physical misalignment. Existing approaches either rely on high-level semantic reasoning without geometric grounding or learn end-to-end policies that lack explicit physical constraints, often resulting in semantically plausible but physically unsafe behaviors. In this paper, we propose MVP-Nav, a physical-aware RGB-only navigation framework that aligns perception, planning, and control with the real 3D world. MVP-Nav reconstructs explicit physical occupancy from monocular observations by leveraging 3D foundation models to project 2D semantic instances into 3D oriented bounding boxes, forming a global spatial semantic representation. To unify high-level semantic reasoning and low-level physical constraints, we introduce a Multi-layer Value Map (MVM) that integrates semantic priorities and reconstructed geometry into a shared cost space, enabling physically grounded geometric planning. Extensive experiments on zero-shot object navigation benchmarks demonstrate that MVP-Nav significantly outperforms existing depth-free methods, achieving state-of-the-art performance and validating that structured physical priors can effectively compensate for the absence of active depth sensors.

📄 PDF Abstract BibTeX arXiv:2606.31919

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Using Deep Reinforcement Learning Methods for Autonomous Vessels in 2D Environments

2020-03-23 · Mohammad Etemad, Nader Zare, Mahtab Sarvmaili, Amilcar Soares 외

Unmanned Surface Vehicles technology (USVs) is an exciting topic that essentially deploys an algorithm to safely and efficiently performs a mission. Although reinforcement learning is a well-known approach to modeling su…

Decision MakingDeep Reinforcement LearningQ-Learningreinforcement-learning+2

Vision-and-Dialog Navigation

2019-07-10 · Jesse Thomason, Michael Murray, Maya Cakmak, Luke Zettlemoyer

Robots navigating in human environments should use language to ask for assistance and be able to understand human responses. To study this challenge, we introduce Cooperative Vision-and-Dialog Navigation, a dataset of ov…

2kVisual Navigation

Continuous Control with Deep Reinforcement Learning for Autonomous Vessels

2021-06-27 · Nader Zare, Bruno Brandoli, Mahtab Sarvmaili, Amilcar Soares 외

Maritime autonomous transportation has played a crucial role in the globalization of the world economy. Deep Reinforcement Learning (DRL) has been applied to automatic path planning to simulate vessel collision avoidance…

Collision Avoidancecontinuous-controlContinuous ControlDeep Reinforcement Learning+3

FeatNavigator: Automatic Feature Augmentation on Tabular Data

2024-06-13 · Jiaming Liang, Chuan Lei, Xiao Qin, Jiani Zhang 외

Data-centric AI focuses on understanding and utilizing high-quality, relevant data in training machine learning (ML) models, thereby increasing the likelihood of producing accurate and useful results. Automatic feature a…

Feature Importance

GPT-4V in Wonderland: Large Multimodal Models for Zero-Shot Smartphone GUI Navigation

2023-11-13 · An Yan, Zhengyuan Yang, Wanrong Zhu, Kevin Lin 외

We present MM-Navigator, a GPT-4V-based agent for the smartphone graphical user interface (GUI) navigation task. MM-Navigator can interact with a smartphone screen as human users, and determine subsequent actions to fulf…

Action Localization