paper-with-me

Papers

Accelerating Reinforcement Learning Agent with EEG-based Implicit Human Feedback

2020-06-30 · Duo Xu, Mohit Agarwal, Ekansh Gupta, Faramarz Fekri, Raghupathy Sivakumar

Providing Reinforcement Learning (RL) agents with human feedback can dramatically improve various aspects of learning. However, previous methods require human observer to give inputs explicitly (e.g., press buttons, voice interface), burdening the human in the loop of RL agent's learning process. Further, it is sometimes difficult or impossible to obtain the explicit human advise (feedback), e.g., autonomous driving, disabled rehabilitation, etc. In this work, we investigate capturing human's intrinsic reactions as implicit (and natural) feedback through EEG in the form of error-related potentials (ErrP), providing a natural and direct way for humans to improve the RL agent learning. As such, the human intelligence can be integrated via implicit feedback with RL algorithms to accelerate the learning of RL agent. We develop three reasonably complex 2D discrete navigational games to experimentally evaluate the overall performance of the proposed work. Major contributions of our work are as follows, (i) we propose and experimentally validate the zero-shot learning of ErrPs, where the ErrPs can be learned for one game, and transferred to other unseen games, (ii) we propose a novel RL framework for integrating implicit human feedbacks via ErrPs with RL agent, improving the label efficiency and robustness to human mistakes, and (iii) compared to prior works, we scale the application of ErrPs to reasonably complex environments, and demonstrate the significance of our approach for accelerated learning through real user experiments.

📄 PDF Abstract BibTeX arXiv:2006.16498

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingEEGElectroencephalogram (EEG)reinforcement-learningReinforcement LearningReinforcement Learning (RL)Zero-Shot Learning

Similar Papers 제목 키워드 기반

Deep Reinforcement Learning with Implicit Human Feedback

2020-01-01 · ICLR 2020 1 · Duo Xu, Mohit Agarwal, Raghupathy Sivakumar, Faramarz Fekri

We consider the following central question in the field of Deep Reinforcement Learning (DRL): How can we use implicit human feedback to accelerate and optimize the training of a DRL algorithm? State-of-the-art methods re…

Atari GamesDeep Reinforcement LearningEEGElectroencephalogram (EEG)+3

Accelerating the Learning of TAMER with Counterfactual Explanations

2021-08-03 · Jakob Karalus, Felix Lindner

The capability to interactively learn from human feedback would enable agents in new settings. For example, even novice users could train service robots in new tasks naturally and interactively. Human-in-the-loop Reinfor…

counterfactualreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm

2025-03-05 · HyeonJun Kim, Kanghoon Lee, Junho Park, Jiachen Li 외

Multi-Agent Reinforcement Learning (MARL) has shown promise in solving complex problems involving cooperation and competition among agents, such as an Unmanned Surface Vehicle (USV) swarm used in search and rescue, surve…

Collision AvoidanceFairnessLanguage ModelingLanguage Modelling+4

Reinforcement Learning from Implicit Neural Feedback for Human-Aligned Robot Control

2025-11-18 · Suzie Kim arxiv

Conventional reinforcement learning (RL) approaches often struggle to learn effective policies under sparse reward conditions, necessitating the manual design of complex, task-specific reward functions. To address this l…

Reinforcement Learning

Towards Reinforcement Learning from Neural Feedback: Mapping fNIRS Signals to Agent Performance

2025-11-17 · Julia Santaniello, Matthew Russell, Benson Jiang, Donatello Sassaroli 외 arxiv

Reinforcement Learning from Human Feedback (RLHF) is a methodology that aligns agent behavior with human preferences by integrating user feedback into the agent's training process. This paper introduces a framework that …

Multi-class ClassificationReinforcement Learning