paper-with-me

홈 › Papers

KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance

2024-08-06 · Jingxian Lu, Wenke Xia, Dong Wang, Zhigang Wang, Bin Zhao, Di Hu, Xuelong Li

Online Imitation Learning struggles with the gap between extensive online exploration space and limited expert trajectories, hindering efficient exploration due to inaccurate reward estimation. Inspired by the findings from cognitive neuroscience, we hypothesize that an agent could estimate precise task-aware reward for efficient online exploration, through decomposing the target task into the objectives of "what to do" and the mechanisms of "how to do". In this work, we introduce the hybrid Key-state guided Online Imitation (KOI) learning method, which leverages the integration of semantic and motion key states as guidance for reward estimation. Initially, we utilize visual-language models to extract semantic key states from expert trajectory, indicating the objectives of "what to do". Within the intervals between semantic key states, optical flow is employed to capture motion key states to understand the mechanisms of "how to do". By integrating a thorough grasp of hybrid key states, we refine the trajectory-matching reward computation, accelerating online imitation learning with task-aware exploration. We evaluate not only the success rate of the tasks in the Meta-World and LIBERO environments, but also the trend of variance during online imitation learning, proving that our method is more sample efficient. We also conduct real-world robotic manipulation experiments to validate the efficacy of our method, demonstrating the practical applicability of our KOI method. Videos and code are available at https://gewu-lab.github.io/Keystate_Online_Imitation/.

📄 PDF Abstract BibTeX arXiv:2408.02912

Code (0)

등록된 구현이 없습니다.

Tasks

Efficient ExplorationImitation LearningOptical Flow Estimation

Similar Papers 제목 키워드 기반

Accelerating Entrepreneurial Decision-Making Through Hybrid Intelligence

2021-05-07 · Dominik Dellermann

Accelerating Entrepreneurial Decision-Making Through Hybrid Intelligence DESIGN PARADIGMS AND PRINCIPLES FOR DECISIONAL GUIDANCE IN ENTREPRENEURSHIP

Decision Making

Target-Driven Distillation: Consistency Distillation with Target Timestep Selection and Decoupled Guidance

2024-09-02 · Cunzheng Wang, Ziyuan Guo, Yuxuan Duan, Huaxia Li 외

Consistency distillation methods have demonstrated significant success in accelerating generative tasks of diffusion models. However, since previous consistency distillation methods use simple and straightforward strateg…

DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance

2025-04-02 · Yuxuan Luo, Zhengkun Rong, Lizhen Wang, Longhao Zhang 외

While recent image-based human animation methods achieve realistic body and facial motion synthesis, critical gaps remain in fine-grained holistic controllability, multi-scale adaptability, and long-term temporal coheren…

Human AnimationImage AnimationMotion Synthesis

MolGuidance: Advanced Guidance Strategies for Conditional Molecular Generation with Flow Matching

2025-12-13 · Jirui Jin, Cheng Zeng, Pawan Prakash, Ellad B. Tadmor 외 arxiv

Key objectives in conditional molecular generation include ensuring chemical validity, aligning generated molecules with target properties, promoting structural diversity, and enabling efficient sampling for discovery. R…

Accelerating recurrent neural network language model based online speech recognition system

2018-01-30 · Kyungmin Lee, Chiyoun Park, Namhoon Kim, Jaewon Lee

This paper presents methods to accelerate recurrent neural network based language models (RNNLMs) for online speech recognition systems. Firstly, a lossy compression of the past hidden layer outputs (history vector) with…

CPUGPULanguage ModelingLanguage Modelling+2