paper-with-me

홈 › Papers

Yell At Your Robot: Improving On-the-Fly from Language Corrections

2024-03-19 · Lucy Xiaoyang Shi, Zheyuan Hu, Tony Z. Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, Chelsea Finn

Hierarchical policies that combine language and low-level control have been shown to perform impressively long-horizon robotic tasks, by leveraging either zero-shot high-level planners like pretrained language and vision-language models (LLMs/VLMs) or models trained on annotated robotic demonstrations. However, for complex and dexterous skills, attaining high success rates on long-horizon tasks still represents a major challenge -- the longer the task is, the more likely it is that some stage will fail. Can humans help the robot to continuously improve its long-horizon task performance through intuitive and natural feedback? In this paper, we make the following observation: high-level policies that index into sufficiently rich and expressive low-level language-conditioned skills can be readily supervised with human feedback in the form of language corrections. We show that even fine-grained corrections, such as small movements ("move a bit to the left"), can be effectively incorporated into high-level policies, and that such corrections can be readily obtained from humans observing the robot and making occasional suggestions. This framework enables robots not only to rapidly adapt to real-time language feedback, but also incorporate this feedback into an iterative training scheme that improves the high-level policy's ability to correct errors in both low-level execution and high-level decision-making purely from verbal feedback. Our evaluation on real hardware shows that this leads to significant performance improvement in long-horizon, dexterous manipulation tasks without the need for any additional teleoperation. Videos and code are available at https://yay-robot.github.io/.

📄 PDF Abstract BibTeX arXiv:2403.12910

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Distilling and Retrieving Generalizable Knowledge for Robot Manipulation via Language Corrections

2023-11-17 · Lihan Zha, Yuchen Cui, Li-Heng Lin, Minae Kwon 외

Today's robot policies exhibit subpar performance when faced with the challenge of generalizing to novel environments. Human corrective feedback is a crucial form of guidance to enable such generalization. However, adapt…

Language ModellingLarge Language ModelRobot Manipulation

Correcting Robot Plans with Natural Language Feedback

2022-04-11 · Pratyusha Sharma, Balakumar Sundaralingam, Valts Blukis, Chris Paxton 외

When humans design cost or goal specifications for robots, they often produce specifications that are ambiguous, underspecified, or beyond planners' ability to solve. In these cases, corrections provide a valuable tool f…

Mind Your POV: Convergence of Articles and Editors Towards Wikipedia's Neutrality Norm

2018-09-18 · Umashanthi Pavalanathan, Xiaochuang Han, Jacob Eisenstein

Wikipedia has a strong norm of writing in a 'neutral point of view' (NPOV). Articles that violate this norm are tagged, and editors are encouraged to make corrections. But the impact of this tagging system has not been q…

ArticlesTime SeriesTime Series Analysis

"No, to the Right" -- Online Language Corrections for Robotic Manipulation via Shared Autonomy

2023-01-06 · Yuchen Cui, Siddharth Karamcheti, Raj Palleti, Nidhya Shivakumar 외

Systems for language-guided human-robot interaction must satisfy two key desiderata for broad adoption: adaptivity and learning efficiency. Unfortunately, existing instruction-following agents cannot adapt, lacking the a…

Instruction Following

TATIC: Task-Aware Temporal Learning for Human Intent Inference from Physical Corrections in Human-Robot Collaboration

2026-03-10 · Jiurun Song, Xiao Liang, Minghui Zheng arxiv

In human-robot collaboration (HRC), robots must adapt online to dynamic task constraints and evolving human intent. While physical corrections provide a natural, low-latency channel for operators to convey motion-level a…

Intent Recognition