paper-with-me

홈 › Papers

Multi-Fidelity Hybrid Reinforcement Learning via Information Gain Maximization

2025-09-18 · Houssem Sifaou, Osvaldo Simeone arxiv

Optimizing a reinforcement learning (RL) policy typically requires extensive interactions with a high-fidelity simulator of the environment, which are often costly or impractical. Offline RL addresses this problem by allowing training from pre-collected data, but its effectiveness is strongly constrained by the size and quality of the dataset. Hybrid offline-online RL leverages both offline data and interactions with a single simulator of the environment. In many real-world scenarios, however, multiple simulators with varying levels of fidelity and computational cost are available. In this work, we study multi-fidelity hybrid RL for policy optimization under a fixed cost budget. We introduce multi-fidelity hybrid RL via information gain maximization (MF-HRL-IGM), a hybrid offline-online RL algorithm that implements fidelity selection based on information gain maximization through a bootstrapping approach. Theoretical analysis establishes the no-regret property of MF-HRL-IGM, while empirical evaluations demonstrate its superior performance compared to existing benchmarks.

📄 PDF Abstract BibTeX arXiv:2509.14848

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningOffline RL

Similar Papers 제목 키워드 기반

Semi-analytical Industrial Cooling System Model for Reinforcement Learning

2022-07-26 · Yuri Chervonyi, Praneet Dutta, Piotr Trochim, Octavian Voicu 외

We present a hybrid industrial cooling system model that embeds analytical solutions within a multi-physics simulation. This model is designed for reinforcement learning (RL) applications and balances simplicity with sim…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Task-driven Semantic Coding via Reinforcement Learning

2021-06-07 · Xin Li, Jun Shi, Zhibo Chen

Task-driven semantic video/image coding has drawn considerable attention with the development of intelligent media applications, such as license plate detection, face detection, and medical diagnosis, which focuses on ma…

Face DetectionLicense Plate DetectionMedical DiagnosisQuantization+3

Active3D: Active High-Fidelity 3D Reconstruction via Hierarchical Uncertainty Quantification

2025-11-25 · Yan Li, Yingzhao Li, Gim Hee Lee arxiv

In this paper, we present an active exploration framework for high-fidelity 3D reconstruction that incrementally builds a multi-level uncertainty space and selects next-best-views through an uncertainty-driven motion pla…

3D Reconstruction

CacheRL:Multi-Turn Tool-Calling Agents via Cached Rollouts and Hybrid Reward

2026-06-12 · Md Amirul Islam, Sumiran Thakur, Huancheng Chen, Su Min Park 외 arxiv

We present CacheRL, a system for training small agent foundation models that achieves 92 percent process accuracy on multi-step tool-calling tasks, approaching GPT-5's 94 percent while requiring 100 times less compute. O…

Reinforcement Learning

When to Trust Your Simulator: Dynamics-Aware Hybrid Offline-and-Online Reinforcement Learning

2022-06-27 · Haoyi Niu, Shubham Sharma, Yiwen Qiu, Ming Li 외

Learning effective reinforcement learning (RL) policies to solve real-world complex tasks can be quite challenging without a high-fidelity simulation environment. In most cases, we are only given imperfect simulators wit…

Offline RLreinforcement-learningReinforcement Learning (RL)