Online Behavior Modification for Expressive User Control of RL-Trained Robots
Reinforcement Learning (RL) is an effective method for robots to learn tasks. However, in typical RL, end-users have little to no control over how the robot does the task after the robot has been deployed. To address this, we introduce the idea of online behavior modification, a paradigm in which users have control over behavior features of a robot in real time as it autonomously completes a task using an RL-trained policy. To show the value of this user-centered formulation for human-robot interaction, we present a behavior diversity based algorithm, Adjustable Control Of RL Dynamics (ACORD), and demonstrate its applicability to online behavior modification in simulation and a user study. In the study (n=23) users adjust the style of paintings as a robot traces a shape autonomously. We compare ACORD to RL and Shared Autonomy (SA), and show ACORD affords user-preferred levels of control and expression, comparable to SA, but with the potential for autonomous execution and robustness of RL.
Code (0)
등록된 구현이 없습니다.
Tasks
DiversityReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Beyond Realism: Learning the Art of Expressive Composition with StickerNet
As a widely used operation in image editing workflows, image composition has traditionally been studied with a focus on achieving visual realism and semantic plausibility. However, in practical editing scenarios of the m…
Image EditingCounterfactual Explanation and Causal Inference in Service of Robustness in Robot Control
We propose an architecture for training generative models of counterfactual conditionals of the form, 'can we modify event A to cause B instead of C?', motivated by applications in robot control. Using an 'adversarial tr…
Causal InferencecounterfactualCounterfactual ExplanationInteractive Constrained MAP-Elites: Analysis and Evaluation of the Expressiveness of the Feature Dimensions
We propose the Interactive Constrained MAP-Elites, a quality-diversity solution for game content generation, implemented as a new feature of the Evolutionary Dungeon Designer: a mixed-initiative co-creativity tool for de…
DiversityHow to "Improve" Prediction Using Behavior Modification
Many internet platforms that collect behavioral big data use it to predict user behavior for internal purposes and for their business customers (e.g., advertisers, insurers, security forces, governments, political consul…
Decision MakingPredictionMulti-Behavior Graph Neural Networks for Recommender System
Recommender systems have been demonstrated to be effective to meet user's personalized interests for many online services (e.g., E-commerce and online advertising platforms). Recent years have witnessed the emerging succ…
Collaborative FilteringGraph Neural NetworkRecommendation SystemsTAG