paper-with-me

홈 › Papers

Adaptive Multi-Fidelity Reinforcement Learning for Variance Reduction in Engineering Design Optimization

2025-03-23 · Akash Agrawal, Christopher McComb

Multi-fidelity Reinforcement Learning (RL) frameworks efficiently utilize computational resources by integrating analysis models of varying accuracy and costs. The prevailing methodologies, characterized by transfer learning, human-inspired strategies, control variate techniques, and adaptive sampling, predominantly depend on a structured hierarchy of models. However, this reliance on a model hierarchy can exacerbate variance in policy learning when the underlying models exhibit heterogeneous error distributions across the design space. To address this challenge, this work proposes a novel adaptive multi-fidelity RL framework, in which multiple heterogeneous, non-hierarchical low-fidelity models are dynamically leveraged alongside a high-fidelity model to efficiently learn a high-fidelity policy. Specifically, low-fidelity policies and their experience data are adaptively used for efficient targeted learning, guided by their alignment with the high-fidelity policy. The effectiveness of the approach is demonstrated in an octocopter design optimization problem, utilizing two low-fidelity models alongside a high-fidelity simulator. The results demonstrate that the proposed approach substantially reduces variance in policy learning, leading to improved convergence and consistent high-quality solutions relative to traditional hierarchical multi-fidelity RL methods. Moreover, the framework eliminates the need for manually tuning model usage schedules, which can otherwise introduce significant computational overhead. This positions the framework as an effective variance-reduction strategy for multi-fidelity RL, while also mitigating the computational and operational burden of manual fidelity scheduling.

📄 PDF Abstract BibTeX arXiv:2503.18229

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)SchedulingTransfer Learning

Similar Papers 제목 키워드 기반

Adaptive Machine Learning-Driven Multi-Fidelity Stratified Sampling for Failure Analysis of Nonlinear Stochastic Systems

2025-08-01 · Liuyun Xu, Seymour M. J. Spence arxiv

Existing variance reduction techniques used in stochastic simulations for rare event analysis still require a substantial number of model evaluations to estimate small failure probabilities. In the context of complex, no…

Multifidelity Reinforcement Learning with Control Variates

2022-06-10 · Sami Khairy, Prasanna Balaprakash

In many computational science and engineering applications, the output of a system of interest corresponding to a given input can be queried at different levels of fidelity with different costs. Typically, low-fidelity d…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Context-aware learning of hierarchies of low-fidelity models for multi-fidelity uncertainty quantification

2022-11-20 · Ionut-Gabriel Farcas, Benjamin Peherstorfer, Tobias Neckel, Frank Jenko 외

Multi-fidelity Monte Carlo methods leverage low-fidelity and surrogate models for variance reduction to make tractable uncertainty quantification even when numerically simulating the physical systems of interest with hig…

Uncertainty Quantification

Variance Reduction on General Adaptive Stochastic Mirror Descent

2020-12-26 · Wenjie Li, Zhanyu Wang, Yichen Zhang, Guang Cheng

In this work, we investigate the idea of variance reduction by studying its properties with general adaptive mirror descent algorithms in nonsmooth nonconvex finite-sum optimization problems. We propose a simple yet gene…

RMFGP: Rotated Multi-fidelity Gaussian process with Dimension Reduction for High-dimensional Uncertainty Quantification

2022-04-11 · Jiahao Zhang, Shiqi Zhang, Guang Lin

Multi-fidelity modelling arises in many situations in computational science and engineering world. It enables accurate inference even when only a small set of accurate data is available. Those data often come from a high…

Active LearningDimensionality ReductionregressionUncertainty Quantification