paper-with-me

홈 › Papers

Non-Differentiable Supervised Learning with Evolution Strategies and Hybrid Methods

2019-06-07 · Karel Lenc, Erich Elsen, Tom Schaul, Karen Simonyan

In this work we show that Evolution Strategies (ES) are a viable method for learning non-differentiable parameters of large supervised models. ES are black-box optimization algorithms that estimate distributions of model parameters; however they have only been used for relatively small problems so far. We show that it is possible to scale ES to more complex tasks and models with millions of parameters. While using ES for differentiable parameters is computationally impractical (although possible), we show that a hybrid approach is practically feasible in the case where the model has both differentiable and non-differentiable parameters. In this approach we use standard gradient-based methods for learning differentiable weights, while using ES for learning non-differentiable parameters - in our case sparsity masks of the weights. This proposed method is surprisingly competitive, and when parallelized over multiple devices has only negligible training time overhead compared to training with gradient descent. Additionally, this method allows to train sparse models from the first training step, so they can be much larger than when using methods that require training dense models first. We present results and analysis of supervised feed-forward models (such as MNIST and CIFAR-10 classification), as well as recurrent models, such as SparseWaveRNN for text-to-speech.

📄 PDF Abstract BibTeX arXiv:1906.03139

Code (0)

등록된 구현이 없습니다.

Tasks

text-to-speechText to Speech

Similar Papers 제목 키워드 기반

Guiding Evolutionary Strategies by Differentiable Robot Simulators

2021-10-01 · Vladislav Kurenkov, Bulat Maksudov

In recent years, Evolutionary Strategies were actively explored in robotic tasks for policy search as they provide a simpler alternative to reinforcement learning algorithms. However, this class of algorithms is often cl…

reinforcement-learningReinforcement Learning (RL)

EvoGrad: Metaheuristics in a Differentiable Wonderland

2025-05-28 · Beatrice F. R. Citterio, Andrea Tangherloni

Differentiable programming has revolutionised optimisation by enabling efficient gradient-based training of complex models, such as Deep Neural Networks (NNs) with billions and trillions of parameters. However, tradition…

Integrating Mechanistic and Data-Driven Models for Neurological Disorders through Differentiable Programming

2026-06-04 · Shah Pallav Dhanendrakumar, Saikat Pal, Sitikantha Roy arxiv

Advances in computational modeling, neuroimaging, and artificial intelligence are revolutionizing the modeling of neurological disorders for improved diagnostics, prognosis, and treatment planning. Mechanistic models pro…

The Hybridization of Branch and Bound with Metaheuristics for Nonconvex Multiobjective Optimization

2022-12-09 · Wei-tian Wu, Xin-min Yang

A hybrid framework combining the branch and bound method with multiobjective evolutionary algorithms is proposed for nonconvex multiobjective optimization. The hybridization exploits the complementary character of the tw…

Evolutionary AlgorithmsMultiobjective Optimization

SpiderNet: Hybrid Differentiable-Evolutionary Architecture Search via Train-Free Metrics

2022-04-20 · Rob Geada, Andrew Stephen McGough

Neural Architecture Search (NAS) algorithms are intended to remove the burden of manual neural network design, and have shown to be capable of designing excellent models for a variety of well-known problems. However, the…

Neural Architecture Search