paper-with-me

홈 › Papers

InterPReT: Interactive Policy Restructuring and Training Enable Effective Imitation Learning from Laypersons

2026-02-04 · Feiyu Gavin Zhu, Jean Oh, Reid Simmons arxiv

Imitation learning has shown success in many tasks by learning from expert demonstrations. However, most existing work relies on large-scale demonstrations from technical professionals and close monitoring of the training process. These are challenging for a layperson when they want to teach the agent new skills. To lower the barrier of teaching AI agents, we propose Interactive Policy Restructuring and Training (InterPReT), which takes user instructions to continually update the policy structure and optimize its parameters to fit user demonstrations. This enables end-users to interactively give instructions and demonstrations, monitor the agent's performance, and review the agent's decision-making strategies. A user study (N=34) on teaching an AI agent to drive in a racing game confirms that our approach yields more robust policies without impairing system usability, compared to a generic imitation learning baseline, when a layperson is responsible for both giving demonstrations and determining when to stop. This shows that our method is more suitable for end-users without much technical background in machine learning to train a dependable policy

📄 PDF Abstract BibTeX arXiv:2602.04213

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Crisp: Cognitive Restructuring of Negative Thoughts through Multi-turn Supportive Dialogues

2025-04-24 · Jinfeng Zhou, Yuxuan Chen, Jianing Yin, Yongkang Huang 외

Cognitive Restructuring (CR) is a psychotherapeutic process aimed at identifying and restructuring an individual's negative thoughts, arising from mental health challenges, into more helpful and positive ones via multi-t…

Sentence

Imagination-Augmented Hierarchical Reinforcement Learning for Safe and Interactive Autonomous Driving in Urban Environments

2023-11-17 · Sang-Hyun Lee, Yoonjae Jung, Seung-Woo Seo

Hierarchical reinforcement learning (HRL) incorporates temporal abstraction into reinforcement learning (RL) by explicitly taking advantage of hierarchical structure. Modern HRL typically designs a hierarchical agent com…

Autonomous DrivingHierarchical Reinforcement LearningReinforcement Learning (RL)

CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning

2025-09-29 · Shijie Zhang, Zheng Xiao, Shiyu Liu, Guohao Sun 외 arxiv

Online reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning abilities of large language models, but most methods still optimize reasoning trajectories over th…

Reinforcement LearningMathematical ReasoningData Augmentation

MUSE: An Interactive Meta-Agent for Understanding and Steering LLM-powered Data Science Systems

2026-08-17 · Wei-Hao Chen, Weixi Tong, Yuan Tian, Chenglong Wang 외 arxiv

Recent advances in large language models have enabled a new class of agentic data science systems that allow users to complete complex data science workflows through natural language. Although these systems can significa…

Towards Integrated Glance To Restructuring in Combinatorial Optimization

2015-12-20 · Mark Sh. Levin

The paper focuses on a new class of combinatorial problems which consists in restructuring of solutions (as sets/structures) in combinatorial optimization. Two main features of the restructuring process are examined: (i)…

ClusteringCombinatorial OptimizationMultiple-choice