paper-with-me

홈 › Papers

Skill Preferences: Learning to Extract and Execute Robotic Skills from Human Feedback

2021-08-11 · Xiaofei Wang, Kimin Lee, Kourosh Hakhamaneshi, Pieter Abbeel, Michael Laskin

A promising approach to solving challenging long-horizon tasks has been to extract behavior priors (skills) by fitting generative models to large offline datasets of demonstrations. However, such generative models inherit the biases of the underlying data and result in poor and unusable skills when trained on imperfect demonstration data. To better align skill extraction with human intent we present Skill Preferences (SkiP), an algorithm that learns a model over human preferences and uses it to extract human-aligned skills from offline data. After extracting human-preferred skills, SkiP also utilizes human feedback to solve down-stream tasks with RL. We show that SkiP enables a simulated kitchen robot to solve complex multi-step manipulation tasks and substantially outperforms prior leading RL algorithms with human preferences as well as leading skill extraction algorithms without human preferences.

📄 PDF Abstract BibTeX arXiv:2108.05382

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Efficient Bimanual Manipulation Using Learned Task Schemas

2019-09-30 · Rohan Chitnis, Shubham Tulsiani, Saurabh Gupta, Abhinav Gupta

We address the problem of effectively composing skills to solve sparse-reward tasks in the real world. Given a set of parameterized skills (such as exerting a force or doing a top grasp at a location), our goal is to lea…

Reinforcement Learning

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents

2026-02-24 · Yanna Jiang, Delong Li, Haiyu Deng, Baihe Ma 외 arxiv

Agentic systems increasingly rely on reusable procedural capabilities, \textit{a.k.a., agentic skills}, to execute long-horizon workflows reliably. These capabilities are callable modules that package procedural knowledg…

WebXSkill: Skill Learning for Autonomous Web Agents

2026-04-14 · Zhaoyang Wang, Qianhui Wu, Xuchao Zhang, Chaoyun Zhang 외 arxiv

Autonomous web agents powered by large language models (LLMs) remain brittle on long-horizon browser workflows. A key bottleneck is a grounding gap in existing skill formulations: textual workflow skills provide natural …

Generalized Animal Imitator: Agile Locomotion with Versatile Motion Prior

2023-10-02 · Ruihan Yang, Zhuoqun Chen, Jianhan Ma, Chongyi Zheng 외

The agility of animals, particularly in complex activities such as running, turning, jumping, and backflipping, stands as an exemplar for robotic system design. Transferring this suite of behaviors to legged robotic syst…

RH20T-P: A Primitive-Level Robotic Dataset Towards Composable Generalization Agents

2024-03-28 · Zeren Chen, Zhelun Shi, Xiaoya Lu, Lehan He 외

Achieving generalizability in solving out-of-distribution tasks is one of the ultimate goals of learning robotic manipulation. Recent progress of Vision-Language Models (VLMs) has shown that VLM-based task planners can a…

Motion Planning