paper-with-me

Papers

DUViN: Diffusion-Based Underwater Visual Navigation via Knowledge-Transferred Depth Features

2025-09-03 · Jinghe Yang, Minh-Quan Le, Mingming Gong, Ye Pu arxiv

Autonomous underwater navigation remains a challenging problem due to limited sensing capabilities and the difficulty of constructing accurate maps in underwater environments. In this paper, we propose a Diffusion-based Underwater Visual Navigation policy via knowledge-transferred depth features, named DUViN, which enables vision-based end-to-end 4-DoF motion control for underwater vehicles in unknown environments. DUViN guides the vehicle to avoid obstacles and maintain a safe and perception awareness altitude relative to the terrain without relying on pre-built maps. To address the difficulty of collecting large-scale underwater navigation datasets, we propose a method that ensures robust generalization under domain shifts from in-air to underwater environments by leveraging depth features and introducing a novel model transfer strategy. Specifically, our training framework consists of two phases: we first train the diffusion-based visual navigation policy on in-air datasets using a pre-trained depth feature extractor. Secondly, we retrain the extractor on an underwater depth estimation task and integrate the adapted extractor into the trained navigation policy from the first step. Experiments in both simulated and real-world underwater environments demonstrate the effectiveness and generalization of our approach. The experimental videos are available at https://www.youtube.com/playlist?list=PLqt2s-RyCf1gfXJgFzKjmwIqYhrP4I-7Y.

📄 PDF Abstract BibTeX arXiv:2509.02983

Code (0)

등록된 구현이 없습니다.

Tasks

Visual NavigationDepth Estimation

Similar Papers 제목 키워드 기반

Deep Learning for Visual Navigation of Underwater Robots

2023-10-30 · M. Sunbeam

This paper aims to briefly survey deep learning methods for visual navigation of underwater robotics. The scope of this paper includes the visual perception of underwater robotics with deep learning methods, the availabl…

Deep LearningImitation Learningreinforcement-learningVisual Navigation

Learning A Physical-aware Diffusion Model Based on Transformer for Underwater Image Enhancement

2024-03-03 · Chen Zhao, Chenyu Dong, Weiling Cai

Underwater visuals undergo various complex degradations, inevitably influencing the efficiency of underwater vision tasks. Recently, diffusion models were employed to underwater image enhancement (UIE) tasks, and gained …

Image EnhancementUIE

CAVE-NAV: VLM-Based Autonomous 3D Navigation in Underwater Cave Environments

2026-08-28 · Zhenqi Wu, Yuanjie Lu, Yisheng Zhang, Miao Yu 외 arxiv

Autonomous navigation in underwater cave environments is essential for search-and-rescue operations, scientific exploration, and emergency egress. Traditional navigation systems commonly depend on dense visual features f…

3d sequential image mosaicing for underwater navigation and mapping

2021-10-04 · E. Nocerino, F. Menna, B. Chemisky, P. Drap

Although fully autonomous mapping methods are becoming more and more common and reliable, still the human operator is regularly employed in many 3D surveying missions. In a number of underwater applications, divers or pi…

validVisual Navigation

CaveSeg: Deep Semantic Segmentation and Scene Parsing for Autonomous Underwater Cave Exploration

2023-09-20 · A. Abdullah, T. Barua, R. Tibbetts, Z. Chen 외

In this paper, we present CaveSeg - the first visual learning pipeline for semantic segmentation and scene parsing for AUV navigation inside underwater caves. We address the problem of scarce annotated training data by p…

Scene ParsingSegmentationSemantic Segmentation