Interactive Groupwise Comparison for Reinforcement Learning from Human Feedback
Reinforcement learning from human feedback (RLHF) has emerged as a key enabling technology for aligning AI behaviour with human preferences. The traditional way to collect data in RLHF is via pairwise comparisons: human raters are asked to indicate which one of two samples they prefer. We present an interactive visualisation that better exploits the human visual ability to compare and explore whole groups of samples. The interface is comprised of two linked views: 1) an exploration view showing a contextual overview of all sampled behaviours organised in a hierarchical clustering structure; and 2) a comparison view displaying two selected groups of behaviours for user queries. Users can efficiently explore large sets of behaviours by iterating between these two views. Additionally, we devised an active learning approach suggesting groups for comparison. As shown by our evaluation in six simulated robotics tasks, our approach increases the final rewards by 69.34%. It leads to lower error rates and better policies. We open-source the code that can be easily integrated into the RLHF training loop, supporting research on human-AI alignment.
Code (0)
등록된 구현이 없습니다.
Tasks
Reinforcement LearningActive LearningSimilar Papers 제목 키워드 기반
RLHF-Blender: A Configurable Interactive Interface for Learning from Diverse Human Feedback
To use reinforcement learning from human feedback (RLHF) in practical applications, it is crucial to learn reward models from diverse sources of human feedback and to consider human factors involved in providing feedback…
Accelerating the Learning of TAMER with Counterfactual Explanations
The capability to interactively learn from human feedback would enable agents in new settings. For example, even novice users could train service robots in new tasks naturally and interactively. Human-in-the-loop Reinfor…
counterfactualreinforcement-learningReinforcement LearningReinforcement Learning (RL)Interactive Reinforcement Learning for Table Balancing Robot
With the development of robotics, the use of robots in daily life is increasing, which has led to the need for anyone to easily train robots to improve robot use. Interactive reinforcement learning(IARL) is a method for …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Deep Reinforcement Learningreinforcement-learning+5Boosting Feedback Efficiency of Interactive Reinforcement Learning by Adaptive Learning from Scores
Interactive reinforcement learning has shown promise in learning complex robotic tasks. However, the process can be human-intensive due to the requirement of a large amount of interactive feedback. This paper presents a …
reinforcement-learningReinforcement LearningDeep Reinforcement Learning with Interactive Feedback in a Human-Robot Environment
Robots are extending their presence in domestic environments every day, being more common to see them carrying out tasks in home scenarios. In the future, robots are expected to increasingly perform more complex tasks an…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)