paper-with-me

Papers

FlashNav: Ultra-Fast Policy Training for Robot Navigation within 20 Seconds

2026-06-14 · Shanze Wang, Yiwei Qian, Xinming Zhang, Jun Xue, Siwei Cheng, Xianghui Wang, Qingyuan Hu, Xiaoyu Shen, Wei Zhang arxiv

Deep reinforcement learning has shown strong potential for robot navigation, but its practical deployment is still limited by the long wall-clock cost of policy training. This paper presents FlashNav, a GPU-first framework for ultra-fast range-based robot navigation training. To the best of our knowledge, FlashNav is the first DRL-based robot navigation framework that reaches seconds-level policy training, with the fastest deployable policy trained in less than 20 seconds. The key idea is to align simulation with the navigation MDP: FlashNav preserves the essential components for velocity-level navigation, including occupancy geometry, range sensing, goal-conditioned control, robot motion dynamics, collision handling, termination, and reset, while removing unnecessary rendering and high-fidelity physical details from the training loop. Built on a batched bitmap simulator and a fully GPU-resident training pipeline with our FastDSAC learner, FlashNav generates massive parallel navigation transitions entirely on GPU. Experiments on TurtleBot2 and Unitree Go2 show that FlashNav achieves a 100\% success-rate below 20 seconds on an RTX 5090 and remains within tens of seconds across desktop GPUs. The learned policies further transfer to physical wheeled and legged robots in static and dynamic indoor scenes, demonstrating that DRL-based navigation can be trained at seconds-level speed while preserving deployable obstacle-avoidance behavior.

📄 PDF Abstract BibTeX arXiv:2606.15846

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningRobot Navigation

Similar Papers 제목 키워드 기반

EmbodiedUS-FS: Fast Slow Intelligence for Ultrasound Robotics

2026-06-21 · Fangzhuo Zhang, Xinyu Wang, Xiao Yang, Jinchang Zhang arxiv

Robotic ultrasound scanning in real clinical environments requires both high-level clinical workflow reasoning and low-level closed-loop execution. Physicians natural-language instructions often contain implicit anatomic…

UltraDP: Generalizable Carotid Ultrasound Scanning with Force-Aware Diffusion Policy

2025-11-19 · Ruoqu Chen, Xiangjie Yan, Kangchen Lv, Gao Huang 외 arxiv

Ultrasound scanning is a critical imaging technique for real-time, non-invasive diagnostics. However, variations in patient anatomy and complex human-in-the-loop interactions pose significant challenges for autonomous ro…

DAISS: Phase-Aware Imitation Learning for Dual-Arm Robotic Ultrasound-Guided Interventions

2026-03-08 · Feng Li, Pei Liu, Shiting Wang, Ning Wang 외 arxiv

Imitation learning has shown strong potential for automating complex robotic manipulation. In medical robotics, ultrasound-guided needle insertion demands precise bimanual coordination, as clinicians must simultaneously …

Coaching a Robotic Sonographer: Learning Robotic Ultrasound with Sparse Expert's Feedback

2024-09-03 · Deepak Raina, Mythra V. Balakuntala, Byung Wook Kim, Juan Wachs 외

Ultrasound is widely employed for clinical intervention and diagnosis, due to its advantages of offering non-invasive, radiation-free, and real-time imaging. However, the accessibility of this dexterous procedure is limi…

UltraDexGrasp: Learning Universal Dexterous Grasping for Bimanual Robots with Synthetic Data

2026-03-05 · Sizhe Yang, Yiman Xie, Zhixuan Liang, Yang Tian 외 arxiv

Grasping is a fundamental capability for robots to interact with the physical world. Humans, equipped with two hands, autonomously select appropriate grasp strategies based on the shape, size, and weight of objects, enab…

Robotic GraspingPoint Clouds