paper-with-me

홈 › Papers

Feasibility-aware Imitation Learning from Observation with Multimodal Feedback

2026-02-17 · Kei Takahashi, Hikaru Sasaki, Takamitsu Matsubara arxiv

Imitation learning frameworks that learn robot control policies from demonstrators' motions via hand-mounted demonstration interfaces have attracted increasing attention. However, due to differences in physical characteristics between demonstrators and robots, this approach faces two limitations: i) the demonstration data do not include robot actions, and ii) the demonstrated motions may be infeasible for robots. These limitations make policy learning difficult. To address them, we propose Feasibility-Aware Behavior Cloning from Observation (FABCO). FABCO integrates behavior cloning from observation, which complements robot actions using robot dynamics models, with feasibility estimation. In feasibility estimation, the demonstrated motions are evaluated using a robot-dynamics model, learned from the robot's execution data, to assess reproducibility under the robot's dynamics. The estimated feasibility is used for multimodal feedback and feasibility-aware policy learning to improve the demonstrator's motions and learn robust policies. Multimodal feedback provides feasibility through the demonstrator's visual and haptic senses to promote feasible demonstrated motions. Feasibility-aware policy learning reduces the influence of demonstrated motions that are infeasible for robots, enabling the learning of policies that robots can execute stably. We conducted experiments with 15 participants on two tasks and confirmed that FABCO improves imitation learning performance by more than 3.2 times compared to the case without feasibility feedback.

📄 PDF Abstract BibTeX arXiv:2602.15351

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Feasibility-aware Imitation Learning from Observations through a Hand-mounted Demonstration Interface

2025-03-12 · Kei Takahashi, Hikaru Sasaki, Takamitsu Matsubara

Imitation learning through a demonstration interface is expected to learn policies for robot automation from intuitive human demonstrations. However, due to the differences in human and robot movement characteristics, a …

Imitation Learning

Exploring Automated Recognition of Instructional Activity and Discourse from Multimodal Classroom Data

2025-11-26 · Ivo Bueno, Ruikun Hou, Babette Bühler, Tim Fütterer 외 arxiv

Observation of classroom interactions can provide concrete feedback to teachers, but current methods rely on manual annotation, which is resource-intensive and hard to scale. This work explores AI-driven analysis of clas…

Surface Constraint Policy for Learning Surface-Constrained and Dynamically Feasible Robot Skills

2026-05-29 · Shuai Ke, Jiexin Zhang, Huan Zhao, Zhiao Wei 외 arxiv

Diffusion-based imitation learning methods have driven rapid progress in robot dexterous manipulation tasks. However, they have limitations when applied to tasks that involve complex free-form surface constraints because…

Can Explicit Physical Feasibility Benefit VLA Learning? An Empirical Study

2026-04-20 · Yubai Wei, Chen Wu, Hashem Haghbayan arxiv

Vision-Language-Action (VLA) models map multimodal inputs directly to robot actions and are typically trained through large-scale imitation learning. While this paradigm has shown strong performance, prevailing VLA train…

RLO-MPC: Robust Learning-Based Output Feedback MPC for Improving the Performance of Uncertain Systems in Iterative Tasks

2021-10-01 · Lukas Brunke, SiQi Zhou, Angela P. Schoellig

In this work we address the problem of performing a repetitive task when we have uncertain observations and dynamics. We formulate this problem as an iterative infinite horizon optimal control problem with output feedbac…

Model Predictive Control