paper-with-me

Papers

Reinforcement Learning for Follow-the-Leader Robotic Endoscopic Navigation via Synthetic Data

2026-01-06 · Sicong Gao, Chen Qian, Laurence Xian, Liao Wu, Maurice Pagnucco, Yang Song arxiv

Autonomous navigation is crucial for both medical and industrial endoscopic robots, enabling safe and efficient exploration of narrow tubular environments without continuous human intervention, where avoiding contact with the inner walls has been a longstanding challenge for prior approaches. We present a follow-the-leader endoscopic robot based on a flexible continuum structure designed to minimize contact between the endoscope body and intestinal walls, thereby reducing patient discomfort. To achieve this objective, we propose a vision-based deep reinforcement learning framework guided by monocular depth estimation. A realistic intestinal simulation environment was constructed in \textit{NVIDIA Omniverse} to train and evaluate autonomous navigation strategies. Furthermore, thousands of synthetic intraluminal images were generated using NVIDIA Replicator to fine-tune the Depth Anything model, enabling dense three-dimensional perception of the intestinal environment with a single monocular camera. Subsequently, we introduce a geometry-aware reward and penalty mechanism to enable accurate lumen tracking. Compared with the original Depth Anything model, our method improves $δ_{1}$ depth accuracy by 39.2% and reduces the navigation J-index by 0.67 relative to the second-best method, demonstrating the robustness and effectiveness of the proposed approach.

📄 PDF Abstract BibTeX arXiv:2601.02798

Code (0)

등록된 구현이 없습니다.

Tasks

Monocular Depth EstimationReinforcement Learning

Similar Papers 제목 키워드 기반

BiliVLA: Scene-Aware Vision-Language-Action Model with Reinforcement Learning for Autonomous Biliary Endoscopic Navigation

2026-06-22 · Jinsong Lin, Chi Kit Ng, Zhiyong Xiong, Zikang Pan 외 arxiv

Endoscopic retrograde cholangiopancreatography (ERCP) demands precise endoscopic navigation and stable biliary cannulation within a narrow monocular field characterized by specular reflections, partial occlusions, and fr…

Reinforcement Learning

Communication-Free Collective Navigation for a Swarm of UAVs via LiDAR-Based Deep Reinforcement Learning

2026-01-20 · Myong-Yol Choi, Hankyoul Ko, Hanse Cho, Changseung Kim 외 arxiv

This paper presents a deep reinforcement learning (DRL) based controller for collective navigation of unmanned aerial vehicle (UAV) swarms in communication-denied environments, enabling robust operation in complex, obsta…

Reinforcement Learning

BASED: Bundle-Adjusting Surgical Endoscopic Dynamic Video Reconstruction using Neural Radiance Fields

2023-09-27 · Shreya Saha, Zekai Liang, Shan Lin, Jingpei Lu 외

Reconstruction of deformable scenes from endoscopic videos is important for many applications such as intraoperative navigation, surgical visual perception, and robotic surgery. It is a foundational requirement for reali…

NeRFVideo Reconstruction

EndoDDC: Learning Sparse to Dense Reconstruction for Endoscopic Robotic Navigation via Diffusion Depth Completion

2026-02-25 · Yinheng Lin, Yiming Huang, Beilei Cui, Long Bai 외 arxiv

Accurate depth estimation plays a critical role in the navigation of endoscopic surgical robots, forming the foundation for 3D reconstruction and safe instrument guidance. Fine-tuning pretrained models heavily relies on …

3D ReconstructionDepth EstimationDepth Completion

Using Reinforcement Learning to Herd a Robotic Swarm to a Target Distribution

2020-06-29 · Zahi M. Kakish, Karthik Elamvazhuthi, Spring Berman

In this paper, we present a reinforcement learning approach to designing a control policy for a "leader" agent that herds a swarm of "follower" agents, via repulsive interactions, as quickly as possible to a target proba…

Q-Learningreinforcement-learningReinforcement Learning (RL)