paper-with-me

홈 › Papers

Human-in-the-Loop Methods for Data-Driven and Reinforcement Learning Systems

2020-08-30 · Vinicius G. Goecks

Recent successes combine reinforcement learning algorithms and deep neural networks, despite reinforcement learning not being widely applied to robotics and real world scenarios. This can be attributed to the fact that current state-of-the-art, end-to-end reinforcement learning approaches still require thousands or millions of data samples to converge to a satisfactory policy and are subject to catastrophic failures during training. Conversely, in real world scenarios and after just a few data samples, humans are able to either provide demonstrations of the task, intervene to prevent catastrophic actions, or simply evaluate if the policy is performing correctly. This research investigates how to integrate these human interaction modalities to the reinforcement learning loop, increasing sample efficiency and enabling real-time reinforcement learning in robotics and real world scenarios. This novel theoretical foundation is called Cycle-of-Learning, a reference to how different human interaction modalities, namely, task demonstration, intervention, and evaluation, are cycled and combined to reinforcement learning algorithms. Results presented in this work show that the reward signal that is learned based upon human interaction accelerates the rate of learning of reinforcement learning algorithms and that learning from a combination of human demonstrations and interventions is faster and more sample efficient when compared to traditional supervised learning algorithms. Finally, Cycle-of-Learning develops an effective transition between policies learned using human demonstrations and interventions to reinforcement learning. The theoretical foundation developed by this research opens new research paths to human-agent teaming scenarios where autonomous agents are able to learn from human teammates and adapt to mission performance metrics in real-time and in real world scenarios.

📄 PDF Abstract BibTeX arXiv:2008.13221

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Data Driven Reward Initialization for Preference based Reinforcement Learning

2023-02-17 · Mudit Verma, Subbarao Kambhampati

Preference-based Reinforcement Learning (PbRL) methods utilize binary feedback from the human in the loop (HiL) over queried trajectory pairs to learn a reward model in an attempt to approximate the human's underlying re…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Task-oriented grasping for dexterous robots using postural synergies and reinforcement learning

2026-02-24 · Dimitrios Dimou, José Santos-Victor, Plinio Moreno arxiv

In this paper, we address the problem of task-oriented grasping for humanoid robots, emphasizing the need to align with human social norms and task-specific objectives. Existing methods, employ a variety of open-loop and…

Reinforcement Learning

Multi-agent reinforcement learning for intent-based service assurance in cellular networks

2022-08-07 · Satheesh K. Perepu, Jean P. Martins, Ricardo Souza S, Kaushik Dey

Recently, intent-based management has received good attention in telecom networks owing to stringent performance requirements for many of the use cases. Several approaches in the literature employ traditional closed-loop…

ManagementMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Value Driven Representation for Human-in-the-Loop Reinforcement Learning

2020-04-02 · Ramtin Keramati, Emma Brunskill

Interactive adaptive systems powered by Reinforcement Learning (RL) have many potential applications, such as intelligent tutoring systems. In such systems there is typically an external human system designer that is cre…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

A Synergistic Framework of Nonlinear Acoustic Computing and Reinforcement Learning for Real-World Human-Robot Interaction

2025-05-04 · Xiaoliang Chen, Xin Yu, Le Chang, Yunhe Huang 외

This paper introduces a novel framework integrating nonlinear acoustic computing and reinforcement learning to enhance advanced human-robot interaction under complex noise and reverberation. Leveraging physically informe…

reinforcement-learningReinforcement Learningspeech-recognitionSpeech Recognition