paper-with-me

Papers

Evolutionary Warm-Starts for Reinforcement Learning in Industrial Continuous Control

2026-03-23 · Tom Maus, Stephan Frank, Tobias Glasmachers arxiv

Reinforcement learning (RL) is still rarely applied in industrial control, partly due to the difficulty of training reliable agents for real-world conditions. This work investigates how evolution strategies can support RL in such settings by introducing a continuous-control adaptation of an industrial sorting benchmark. The CMA-ES algorithm is used to generate high-quality demonstrations that warm-start RL agents. Results show that CMA-ES-guided initialization significantly improves stability and performance. Furthermore, the demonstration trajectories generated with the CMA-ES provide a strong oracle reference performance level, which is of interest in its own right. The study delivers a focused proof of concept for hybrid evolutionary-RL approaches and a basis for future, more complex industrial applications.

📄 PDF Abstract BibTeX arXiv:2603.26750

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningContinuous Control

Similar Papers 제목 키워드 기반

Linear Complementarity for Regularized Policy Evaluation and Improvement

2010-12-01 · NeurIPS 2010 12 · Jeffrey Johns, Christopher Painter-Wakefield, Ronald Parr

Recent work in reinforcement learning has emphasized the power of L1 regularization to perform feature selection and prevent overfitting. We propose formulating the L1 regularized linear fixed point problem as a linear c…

feature selectionReinforcement LearningReinforcement Learning (RL)

A Feature-Based Comparison of Evolutionary Computing Techniques for Constrained Continuous Optimisation

2015-09-23 · Shayan Poursoltan, Frank Neumann

Evolutionary algorithms have been frequently applied to constrained continuous optimisation problems. We carry out feature based comparisons of different types of evolutionary algorithms such as evolution strategies, dif…

Evolutionary Algorithms

Batch Reinforcement Learning on the Industrial Benchmark: First Experiences

2017-05-20 · Daniel Hein, Steffen Udluft, Michel Tokic, Alexander Hentschel 외

The Particle Swarm Optimization Policy (PSO-P) has been recently introduced and proven to produce remarkable results on interacting with academic reinforcement learning benchmarks in an off-policy, batch-based setting. T…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Continuous Versatile Jumping Using Learned Action Residuals

2023-04-17 · Yuxiang Yang, Xiangyun Meng, Wenhao Yu, Tingnan Zhang 외

Jumping is essential for legged robots to traverse through difficult terrains. In this work, we propose a hierarchical framework that combines optimal control and reinforcement learning to learn continuous jumping motion…

On Self-Adaptive Mutation Restarts for Evolutionary Robotics with Real Rotorcraft

2017-03-31 · Gerard David Howard

Self-adaptive parameters are increasingly used in the field of Evolutionary Robotics, as they allow key evolutionary rates to vary autonomously in a context-sensitive manner throughout the optimisation process. A signifi…