paper-with-me

Papers

VFP: Variational Flow-Matching Policy for Multi-Modal Robot Manipulation

2025-08-03 · Xuanran Zhai, Qianyou Zhao, Qiaojun Yu, Ce Hao arxiv

Flow-matching-based policies have recently emerged as a promising approach for learning-based robot manipulation, offering significant acceleration in action sampling compared to diffusion-based policies. However, conventional flow-matching methods struggle with multi-modality, often collapsing to averaged or ambiguous behaviors in complex manipulation tasks. To address this, we propose the Variational Flow-Matching Policy (VFP), which introduces a variational latent prior for mode-aware action generation and effectively captures both task-level and trajectory-level multi-modality. VFP further incorporates Kantorovich Optimal Transport (K-OT) for distribution-level alignment and utilizes a Mixture-of-Experts (MoE) decoder for mode specialization and efficient inference. We comprehensively evaluate VFP on 41 simulated tasks and 3 real-robot tasks, demonstrating its effectiveness and sampling efficiency in both simulated and real-world settings. Results show that VFP achieves a 49% relative improvement in task success rate over standard flow-based baselines in simulation, and further outperforms them on real-robot tasks, while still maintaining fast inference and a compact model size. More details are available on our project page: https://sites.google.com/view/varfp/

📄 PDF Abstract BibTeX arXiv:2508.01622

Code (0)

등록된 구현이 없습니다.

Tasks

Robot Manipulation

Similar Papers 제목 키워드 기반

Variational Rectified Flow Matching

2025-02-13 · Pengsheng Guo, Alexander G. Schwing

We study Variational Rectified Flow Matching, a framework that enhances classic rectified flow matching by modeling multi-modal velocity vector-fields. At inference time, classic rectified flow matching 'moves' samples f…

Reinforcement Learning for Flow-Matching Policies with Density Transport

2026-06-07 · Boshu Lei, Kostas Daniilidis, Antonio Loquercio arxiv

We present an online reinforcement learning (RL) algorithm for fine-tuning flow-matching policies in continuous-control problems. Our key insight is to view RL-based policy improvement as a transport of action densities …

Reinforcement LearningRobot Manipulation

Streaming Flow Policy: Simplifying diffusion$/$flow-matching policies by treating action trajectories as flow trajectories

2025-05-28 · Sunshine Jiang, Xiaolin Fang, Nicholas Roy, Tomás Lozano-Pérez 외

Recent advances in diffusion$/$flow-matching policies have enabled imitation learning of complex, multi-modal action trajectories. However, they are computationally expensive because they sample a trajectory of trajector…

Imitation Learning

Show-o2: Improved Native Unified Multimodal Models

2025-06-18 · Jinheng Xie, Zhenheng Yang, Mike Zheng Shou

This paper presents improved native unified multimodal models, \emph{i.e.,} Show-o2, that leverage autoregressive modeling and flow matching. Built upon a 3D causal variational autoencoder space, unified visual represent…

Language ModelingLanguage ModellingVideo Generation

LAFP: Preserving Latent Action Structure in Latent Policy Learning via Flow Matching

2026-06-09 · Jiexi Lyu, Xizhou Bu, Qingqiu Huang, Chufeng Tang 외 arxiv

Learning high-quality latent actions from large-scale unlabeled videos, coupled with limited real-world interaction data for training an action decoder, has emerged as a promising paradigm for scalable latent policy lear…