paper-with-me

홈 › Papers

CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model

2025-08-14 · Zhuoyuan Yu, Yuxing Long, Zihan Yang, Chengyan Zeng, Hongwei Fan, Jiyao Zhang, Hao Dong arxiv

Existing vision-and-language navigation models often deviate from the correct trajectory when executing instructions. However, these models lack effective error correction capability, hindering their recovery from errors. To address this challenge, we propose Self-correction Flywheel, a novel post-training paradigm. Instead of considering the model's error trajectories on the training set as a drawback, our paradigm emphasizes their significance as a valuable data source. We have developed a method to identify deviations in these error trajectories and devised innovative techniques to automatically generate self-correction data for perception and action. These self-correction data serve as fuel to power the model's continued training. The brilliance of our paradigm is revealed when we re-evaluate the model on the training set, uncovering new error trajectories. At this time, the self-correction flywheel begins to spin. Through multiple flywheel iterations, we progressively enhance our monocular RGB-based VLA navigation model CorrectNav. Experiments on R2R-CE and RxR-CE benchmarks show CorrectNav achieves new state-of-the-art success rates of 65.1% and 69.3%, surpassing prior best VLA navigation models by 8.2% and 16.4%. Real robot tests in various indoor and outdoor environments demonstrate \method's superior capability of error correction, dynamic obstacle avoidance, and long instruction following.

📄 PDF Abstract BibTeX arXiv:2508.10416

Code (0)

등록된 구현이 없습니다.

Tasks

Instruction Following

Similar Papers 제목 키워드 기반

PAG: Multi-Turn Reinforced LLM Self-Correction with Policy as Generative Verifier

2025-06-12 · Yuhua Jiang, Yuwen Xiong, Yufeng Yuan, Chao Xin 외

Large Language Models (LLMs) have demonstrated impressive capabilities in complex reasoning tasks, yet they still struggle to reliably verify the correctness of their own outputs. Existing solutions to this verification …

Reinforcement Learning (RL)

Analysis and optimization of a novel energy storage flywheel for improved energy capacity

2022-02-20 · Xiaojun Li, Alan Palazzolo

Kinetic/Flywheel energy storage systems (FESS) have re-emerged as a vital technology in many areas such as smart grid, renewable energy, electric vehicle, and high-power applications. FESSs are designed and optimized to …

BlueLM-GUI Technical Report: A Real-Device-Centric Flywheel for Self-Improving Mobile GUI Agents

2026-09-11 · Tong Ye, Kunyang Han, Guozhi Wang, Longqiang Luo 외 arxiv

Mobile GUI agents are shifting from multi-module frameworks to native models trained end-to-end, yet industrial deployment faces three persistent gaps. Sandbox training produces a distribution mismatch with production en…

Reinforcement Learning

DexFlyWheel: A Scalable and Self-improving Data Generation Framework for Dexterous Manipulation

2025-09-28 · Kefei Zhu, Fengshuo Bai, YuanHao Xiang, Yishuai Cai 외 arxiv

Dexterous manipulation is critical for advancing robot capabilities in real-world applications, yet diverse and high-quality datasets remain scarce. Existing data collection methods either rely on human teleoperation or …

Reinforcement LearningData Augmentation

Bootstrapping Language-Guided Navigation Learning with Self-Refining Data Flywheel

2024-12-11 · Zun Wang, Jialu Li, Yicong Hong, Songze Li 외

Creating high-quality data for training robust language-instructed agents is a long-lasting challenge in embodied AI. In this paper, we introduce a Self-Refining Data Flywheel (SRDF) that generates high-quality and large…