paper-with-me

홈 › Papers

Open Problems and Modern Solutions for Deep Reinforcement Learning

2023-02-05 · Weiqin Chen

Deep Reinforcement Learning (DRL) has achieved great success in solving complicated decision-making problems. Despite the successes, DRL is frequently criticized for many reasons, e.g., data inefficient, inflexible and intractable reward design. In this paper, we review two publications that investigate the mentioned issues of DRL and propose effective solutions. One designs the reward for human-robot collaboration by combining the manually designed extrinsic reward with a parameterized intrinsic reward function via the deterministic policy gradient, which improves the task performance and guarantees a stronger obstacle avoidance. The other one applies selective attention and particle filters to rapidly and flexibly attend to and select crucial pre-learned features for DRL using approximate inference instead of backpropagation, thereby improving the efficiency and flexibility of DRL. Potential avenues for future work in both domains are discussed in this paper.

📄 PDF Abstract BibTeX arXiv:2302.02298

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

2020-05-04 · Sergey Levine, Aviral Kumar, George Tucker, Justin Fu

In this tutorial article, we aim to provide the reader with the conceptual tools needed to get started on research on offline reinforcement learning algorithms: reinforcement learning algorithms that utilize previously c…

Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement Learning+1

OR-Gym: A Reinforcement Learning Library for Operations Research Problems

2020-08-14 · Christian D. Hubbs, Hector D. Perez, Owais Sarwar, Nikolaos V. Sahinidis 외

Reinforcement learning (RL) has been widely applied to game-playing and surpassed the best human-level performance in many domains, yet there are few use-cases in industrial or commercial settings. We introduce OR-Gym, a…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Test-time Recursive Thinking: Self-Improvement without External Feedback

2026-02-03 · Yufan Zhuang, Chandan Singh, Liyuan Liu, Yelong Shen 외 arxiv

Modern Large Language Models (LLMs) have shown rapid improvements in reasoning capabilities, driven largely by reinforcement learning (RL) with verifiable rewards. Here, we ask whether these LLMs can self-improve without…

Reinforcement Learning

Learning Algorithmic Solutions to Symbolic Planning Tasks with a Neural Computer

2019-09-25 · Daniel Tanneberg, Elmar Rueckert, Jan Peters

A key feature of intelligent behavior is the ability to learn abstract strategies that transfer to unfamiliar problems. Therefore, we present a novel architecture, based on memory-augmented networks, that is inspired by …

reinforcement-learningReinforcement Learning (RL)Sokoban

Learning Algorithmic Solutions to Symbolic Planning Tasks with a Neural Computer Architecture

2019-10-30 · Daniel Tanneberg, Elmar Rueckert, Jan Peters

A key feature of intelligent behavior is the ability to learn abstract strategies that transfer to unfamiliar problems. Therefore, we present a novel architecture, based on memory-augmented networks, that is inspired by …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Sokoban