paper-with-me

홈 › Papers

Importance mixing: Improving sample reuse in evolutionary policy search methods

2018-08-17 · Aloïs Pourchot, Nicolas Perrin, Olivier Sigaud

Deep neuroevolution, that is evolutionary policy search methods based on deep neural networks, have recently emerged as a competitor to deep reinforcement learning algorithms due to their better parallelization capabilities. However, these methods still suffer from a far worse sample efficiency. In this paper we investigate whether a mechanism known as "importance mixing" can significantly improve their sample efficiency. We provide a didactic presentation of importance mixing and we explain how it can be extended to reuse more samples. Then, from an empirical comparison based on a simple benchmark, we show that, though it actually provides better sample efficiency, it is still far from the sample efficiency of deep reinforcement learning, though it is more stable.

📄 PDF Abstract BibTeX arXiv:1808.05832

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

On the Reuse Bias in Off-Policy Reinforcement Learning

2022-09-15 · Chengyang Ying, Zhongkai Hao, Xinning Zhou, Hang Su 외

Importance sampling (IS) is a popular technique in off-policy evaluation, which re-weights the return of trajectories in the replay buffer to boost sample efficiency. However, training with IS can be unstable and previou…

continuous-controlContinuous ControlMuJoCoOff-policy evaluation+3

Sample Reuse via Importance Sampling in Information Geometric Optimization

2018-05-31 · Shinichi Shirakawa, Youhei Akimoto, Kazuki Ouchi, Kouzou Ohara

In this paper we propose a technique to reduce the number of function evaluations, which is often the bottleneck of the black-box optimization, in the information geometric optimization (IGO) that is a generic framework …

Evolutionary AlgorithmsIncremental Learning

Lifetime policy reuse and the importance of task capacity

2021-06-03 · David M. Bossens, Adam J. Sobey

A long-standing challenge in artificial intelligence is lifelong reinforcement learning, where learners are given many tasks in sequence and must transfer knowledge between tasks while avoiding catastrophic forgetting. P…

reinforcement-learningReinforcement Learning

Variance Reduction based Experience Replay for Policy Optimization

2022-08-25 · Hua Zheng, Wei Xie, M. Ben Feng

For reinforcement learning on complex stochastic systems where many factors dynamically impact the output trajectories, it is desirable to effectively leverage the information from historical samples collected in previou…

Reinforcement Learning (RL)

Sample Dropout: A Simple yet Effective Variance Reduction Technique in Deep Policy Optimization

2023-02-05 · Zichuan Lin, Xiapeng Wu, Mingfei Sun, Deheng Ye 외

Recent success in Deep Reinforcement Learning (DRL) methods has shown that policy optimization with respect to an off-policy distribution via importance sampling is effective for sample reuse. In this paper, we show that…

Deep Reinforcement LearningMuJoCo