paper-with-me

홈 › Papers

FlowDAgger: Human-in-the-Loop Adaptation of Generative Robot Policies in Latent Space

2026-07-09 · Michael Murray, Daphne Chen, Simran Bagaria, Dean Fortier, Tess Hellebrekers, Galen Mullins, Harshavardhan Gajarla, Oier Mees, Maya Cakmak, Andrey Kolobov arxiv

Pretrained generative robot policies based on flow matching and diffusion have achieved impressive results across a wide range of manipulation tasks. Yet real-world deployments routinely expose failure modes outside the pretraining distribution. Closing these gaps typically requires large-scale data collection or online reinforcement learning on physical hardware, which is impractical for rapid and safe adaptation. We present FlowDAgger, a sample- and compute-efficient method for adapting frozen generative robot policies from human interventions in latent space. Our key idea is action inversion: each human expert action is mapped to the noise that would have produced it under the frozen base policy, using reverse-time integration followed by local refinement. The resulting inverted noise provides supervision for a lightweight latent policy that steers the base model at deployment time, enabling rapid skill acquisition while preserving its behavioral priors. We evaluate FlowDAgger in simulation and on real-world bimanual and single-arm manipulation, adapting both action-head VLAs and world-action models from a handful of interventions. FlowDAgger outperforms supervised fine-tuning and latent-space RL baselines and preserves pretrained skills on held-out tasks, offering a practical path for adapting robot foundation models in the real world. Website: https://microsoft.github.io/FlowDAgger

📄 PDF Abstract BibTeX arXiv:2607.08877

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

FlowCorrect: Efficient Interactive Correction of Generative Flow Policies for Robotic Manipulation

2026-02-25 · Edgar Welte, Yitian Shi, Rosa Wolf, Maximillian Gilles 외 arxiv

Generative manipulation policies can fail catastrophically under deployment-time distribution shift, yet many failures are near-misses: the robot reaches almost-correct poses and would succeed with a small corrective mot…

Exploring vestibulo-ocular adaptation in a closed-loop neuro-robotic experiment using STDP. A simulation study

2020-03-03 · Francisco Naveros, Jesus A. Garrido, Angelo Arleo, Eduardo Ros 외

Studying and understanding the computational primitives of our neural system requires for a diverse and complementary set of techniques. In this work, we use the Neuro-robotic Platform (NRP)to evaluate the vestibulo ocul…

Neuroadaptation in Physical Human-Robot Collaboration

2023-09-30 · Avinash Singh, Dikai Liu, Chin-Teng Lin

Robots for physical Human-Robot Collaboration (pHRC) systems need to change their behavior and how they operate in consideration of several factors, such as the performance and intention of a human co-worker and the capa…

Collision AvoidanceEEGElectroencephalogram (EEG)

Human-in-the-loop Optimisation in Robot-assisted Gait Training

2025-10-07 · Andreas Christou, Andreas Sochopoulos, Elliot Lister, Sethu Vijayakumar arxiv

Wearable robots offer a promising solution for quantitatively monitoring gait and providing systematic, adaptive assistance to promote patient independence and improve gait. However, due to significant interpersonal and …

Maximal Adaptation, Minimal Guidance: Permissive Reactive Robot Task Planning with Humans in the Loop

2025-10-14 · Oz Gitelson, Satya Prakash Nayak, Ritam Raha, Anne-Kathrin Schmuck arxiv

We present a novel framework for human-robot \emph{logical} interaction that enables robots to reliably satisfy (infinite horizon) temporal logic tasks while effectively collaborating with humans who pursue independent a…

Robot Task Planning