paper-with-me

홈 › Papers

On the Interaction of Adaptive Population Control with Cumulative Step-Size Adaptation

2024-10-01 · Amir Omeradzic, Hans-Georg Beyer

Three state-of-the-art adaptive population control strategies (PCS) are theoretically and empirically investigated for a multi-recombinative, cumulative step-size adaptation Evolution Strategy $(\mu/\mu_I, \lambda)$-CSA-ES. First, scaling properties for the generation number and mutation strength rescaling are derived on the sphere in the limit of large population sizes. Then, the adaptation properties of three standard CSA-variants are studied as a function of the population size and dimensionality, and compared to the predicted scaling results. Thereafter, three PCS are implemented along the CSA-ES and studied on a test bed of sphere, random, and Rastrigin functions. The CSA-adaptation properties significantly influence the performance of the PCS, which is shown in more detail. Given the test bed, well-performing parameter sets (in terms of scaling, efficiency, and success rate) for both the CSA- and PCS-subroutines are identified.

📄 PDF Abstract BibTeX arXiv:2410.00595

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning

2026-05-13 · Timofey Tomashevskiy arxiv

Safety in reinforcement learning is often specified through cumulative cost constraints, but these trajectory-level guarantees do not directly prevent unsafe individual decisions, especially under nonstationarity. In con…

Reinforcement Learning

The Ratchet Effect in Silico: How Interaction Drives Cumulative Intelligence in Large Language Models

2025-07-25 · Ren Zhuang arxiv

Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic drift. Large language model training, by contrast, still depends primarily on st…

Mathematical Reasoning

Calibration of P-values for calibration and for deviation of a subpopulation from the full population

2022-01-31 · Mark Tygert

The author's recent research papers, "Cumulative deviation of a subpopulation from the full population" and "A graphical method of cumulative differences between two subpopulations" (both published in volume 8 of Springe…

Mathematical Proofs

Privacy Preserving Reinforcement Learning for Population Processes

2024-06-25 · Samuel Yang-Zhao, Kee Siong Ng

We consider the problem of privacy protection in Reinforcement Learning (RL) algorithms that operate over population processes, a practical but understudied setting that includes, for example, the control of epidemics in…

Privacy Preservingreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Context-Aware Optimization of Follow-Up Intervals for Type 2 Diabetes Care Using Markov Decision Processes

2026-06-17 · Parisa Lotfibagha, Kristen Miller, William J. Gallagher, Elizabeth B. Selden 외 arxiv

Chronic disease management relies on regular patient-provider interactions to follow-up on disease progression and control. For Type 2 Diabetes (T2D), current guidelines prescribe fixed time intervals between subsequent …

Dimensionality Reduction