paper-with-me

홈 › Papers

Fast Visuomotor Policy for Robotic Manipulation

2025-10-14 · Jingkai Jia, Tong Yang, Xueyao Chen, Chenhuan Liu, Wenqiang Zhang arxiv

We present a fast and effective policy framework for robotic manipulation, named Energy Policy, designed for high-frequency robotic tasks and resource-constrained systems. Unlike existing robotic policies, Energy Policy natively predicts multimodal actions in a single forward pass, enabling high-precision manipulation at high speed. The framework is built upon two core components. First, we adopt the energy score as the learning objective to facilitate multimodal action modeling. Second, we introduce an energy MLP to implement the proposed objective while keeping the architecture simple and efficient. We conduct comprehensive experiments in both simulated environments and real-world robotic tasks to evaluate the effectiveness of Energy Policy. The results show that Energy Policy matches or surpasses the performance of state-of-the-art manipulation methods while significantly reducing computational overhead. Notably, on the MimicGen benchmark, Energy Policy achieves superior performance with at a faster inference compared to existing approaches.

📄 PDF Abstract BibTeX arXiv:2510.12483

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FreqPolicy: Frequency Autoregressive Visuomotor Policy with Continuous Tokens

2025-06-02 · Yiming Zhong, Yumeng Liu, Chuyang Xiao, Zemin Yang 외

Learning effective visuomotor policies for robotic manipulation is challenging, as it requires generating precise actions while maintaining computational efficiency. Existing methods remain unsatisfactory due to inherent…

Computational Efficiency

FreqPolicy: Efficient Flow-based Visuomotor Policy via Frequency Consistency

2025-06-10 · Yifei Su, Ning Liu, Dong Chen, Zhen Zhao 외

Generative modeling-based visuomotor policies have been widely adopted in robotic manipulation attributed to their ability to model multimodal action distributions. However, the high inference cost of multi-step sampling…

Action GenerationImage GenerationVision-Language-Action

Normalizing Flows are Capable Models for Bi-manual Visuomotor Policy

2025-09-25 · Jialong Li, Simon Kristoffersson Lind, Wenrui Xie, Maj Stenmark 외 arxiv

The field of general-purpose robotics has recently embraced powerful probabilistic diffusion-based models to learn the complex embodiment behaviours. However, existing models often come with significant trade-offs, namel…

SeFA-Policy: Fast and Accurate Visuomotor Policy Learning with Selective Flow Alignment

2025-11-11 · Rong Xue, Jiageng Mao, Mingtong Zhang, Yue Wang arxiv

Developing efficient and accurate visuomotor policies poses a central challenge in robotic imitation learning. While recent rectified flow approaches have advanced visuomotor policy learning, they suffer from a key limit…

StereoPolicy: Improving Robotic Manipulation Policies via Stereo Perception

2026-05-11 · Evans Han, Yunfan Jiang, Yingke Wang, Haoyue Xiao 외 arxiv

Recent advances in robot imitation learning have produced powerful visuomotor policies that manipulate diverse objects from visual inputs. However, monocular observations lack depth information, which is critical for pre…

Point Clouds