paper-with-me

Papers Model-based Reinforcement Learning

“Model-based Reinforcement Learning” 태그가 달린 논문 708편 · 필터 해제

TransDreamerV3: Implanting Transformer In DreamerV3

2025-06-20 · Shruti Sadanand Dongare, Amun Kharel, Jonathan Samuel, Xiaona Zhou

This paper introduces TransDreamerV3, a reinforcement learning model that enhances the DreamerV3 architecture by integrating a transformer encoder. The model is designed to improve memory and decision-making capabilities…

Decision MakingMinecraftModel-based Reinforcement Learningreinforcement-learning+1

On Quantum BSDE Solver for High-Dimensional Parabolic PDEs

2025-06-17 · Howard Su, Huan-Hsin Tseng

We propose a quantum machine learning framework for approximating solutions to high-dimensional parabolic partial differential equations (PDEs) that can be reformulated as backward stochastic differential equations (BSDE…

Model-based Reinforcement LearningQuantum Machine Learning

Relative Entropy Regularized Reinforcement Learning for Efficient Encrypted Policy Synthesis

2025-06-14 · Jihoon Suh, Yeongjun Jang, Kaoru Teranishi, Takashi Tanaka

We propose an efficient encrypted policy synthesis to develop privacy-preserving model-based reinforcement learning. We first demonstrate that the relative-entropy-regularized reinforcement learning framework offers a co…

Model-based Reinforcement LearningPrivacy PreservingQuantizationreinforcement-learning+1

Accelerating Model-Based Reinforcement Learning using Non-Linear Trajectory Optimization

2025-06-03 · Marco Calì, Giulio Giacomuzzo, Ruggero Carli, Alberto Dalla Libera

This paper addresses the slow policy optimization convergence of Monte Carlo Probabilistic Inference for Learning Control (MC-PILCO), a state-of-the-art model-based reinforcement learning (MBRL) algorithm, by integrating…

Model-based Reinforcement Learning

Bregman Centroid Guided Cross-Entropy Method

2025-06-02 · Yuliang Gu, Hongpeng Cao, Marco Caccamo, Naira Hovakimyan

The Cross-Entropy Method (CEM) is a widely adopted trajectory optimizer in model-based reinforcement learning (MBRL), but its unimodal sampling strategy often leads to premature convergence in multimodal landscapes. In t…

DiversityModel-based Reinforcement Learning

World Models for Cognitive Agents: Transforming Edge Intelligence in Future Networks

2025-05-31 · Changyuan Zhao, Ruichen Zhang, Jiacheng Wang, Gaosheng Zhao 외

World models are emerging as a transformative paradigm in artificial intelligence, enabling agents to construct internal representations of their environments for predictive reasoning, planning, and decision-making. By l…

Decision MakingModel-based Reinforcement LearningTrajectory Planning

Calibrated Value-Aware Model Learning with Stochastic Environment Models

2025-05-28 · Claas Voelcker, Anastasiia Pedan, Arash Ahmadian, Romina Abachi 외

The idea of value-aware model learning, that models should produce accurate value estimates, has gained prominence in model-based reinforcement learning. The MuZero loss, which penalizes a model's value function predicti…

Model-based Reinforcement Learning

JEDI: Latent End-to-end Diffusion Mitigates Agent-Human Performance Asymmetry in Model-Based Reinforcement Learning

2025-05-26 · Jing Yu Lim, Zarif Ikram, Samson Yu, Haozhe Ma 외

Recent advances in model-based reinforcement learning (MBRL) have achieved super-human level performance on the Atari100k benchmark, driven by reinforcement learning agents trained on powerful diffusion world models. How…

Model-based Reinforcement Learning

MedDreamer: Model-Based Reinforcement Learning with Latent Imagination on Complex EHRs for Clinical Decision Support

2025-05-26 · Qianyi Xu, Gousia Habib, Dilruk Perera, Mengling Feng

Timely and personalized treatment decisions are essential across a wide range of healthcare settings where patient responses vary significantly and evolve over time. Clinical data used to support these decisions are ofte…

ImputationModel-based Reinforcement LearningRecommendation SystemsReinforcement Learning (RL)

Deep Active Inference Agents for Delayed and Long-Horizon Environments

2025-05-26 · Yavar Taheri Yeganeh, Mohsen Jafari, Andrea Matta

With the recent success of world-model agents, which extend the core idea of model-based reinforcement learning by learning a differentiable model for sample-efficient control across diverse tasks, active inference (AIF)…

Model-based Reinforcement Learning

Raw2Drive: Reinforcement Learning with Aligned World Models for End-to-End Autonomous Driving (in CARLA v2)

2025-05-22 · Zhenjie Yang, Xiaosong Jia, QiFeng Li, Xue Yang 외

Reinforcement Learning (RL) can mitigate the causal confusion and distribution shift inherent to imitation learning (IL). However, applying RL to end-to-end autonomous driving (E2E-AD) remains an open problem for its tra…

Autonomous DrivingBench2DriveCARLA Leaderboard 2.0Imitation Learning+4

Gaze Into the Abyss -- Planning to Seek Entropy When Reward is Scarce

2025-05-22 · Ashish Sundar, Chunbo Luo, Xiaoyang Wang

Model-based reinforcement learning (MBRL) offers an intuitive way to increase the sample efficiency of model-free RL methods by simultaneously training a world model that learns to predict the future. MBRL methods have p…

Model-based Reinforcement LearningModel Predictive Control

Improving planning and MBRL with temporally-extended actions

2025-05-21 · Palash Chatterjee, Roni Khardon

Continuous time systems are often modeled using discrete time dynamics but this requires a small simulation step to maintain accuracy. In turn, this requires a large planning horizon which leads to computationally demand…

Model-based Reinforcement Learning

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning

2025-05-19 · Dongsu Lee, Minhae Kwon

The goal of offline reinforcement learning (RL) is to extract a high-performance policy from the fixed datasets, minimizing performance degradation due to out-of-distribution (OOD) samples. Offline model-based RL (MBRL) …

D4RLModel-based Reinforcement LearningReinforcement Learning (RL)

Policy-Driven World Model Adaptation for Robust Offline Model-based Reinforcement Learning

2025-05-19 · Jiayu Chen, Aravind Venugopal, Jeff Schneider

Offline reinforcement learning (RL) offers a powerful paradigm for data-driven control. Compared to model-free approaches, offline model-based RL (MBRL) explicitly learns a world model from a static dataset and uses it a…

D4RLmodelModel-based Reinforcement LearningMuJoCo+1

Multi-Goal Dexterous Hand Manipulation using Probabilistic Model-based Reinforcement Learning

2025-04-30 · Yingzhuo Jiang, Wenjun Huang, Rongdun Lin, Chenyang Miao 외

This paper tackles the challenge of learning multi-goal dexterous hand manipulation tasks using model-based Reinforcement Learning. We propose Goal-Conditioned Probabilistic Model Predictive Control (GC-PMPC) by designin…

Model-based Reinforcement LearningModel Predictive Control

PIN-WM: Learning Physics-INformed World Models for Non-Prehensile Manipulation

2025-04-23 · Wenxuan Li, Hang Zhao, Zhiyuan Yu, Yu Du 외

While non-prehensile manipulation (e.g., controlled pushing/poking) constitutes a foundational robotic skill, its learning remains challenging due to the high sensitivity to complex physical interactions involving fricti…

FrictionModel-based Reinforcement LearningState Estimation

Data-Assimilated Model-Based Reinforcement Learning for Partially Observed Chaotic Flows

2025-04-23 · Defne E. Ozan, Andrea Nóvoa, Luca Magri

The goal of many applications in energy and transport sectors is to control turbulent flows. However, because of chaotic dynamics and high dimensionality, the control of turbulent flows is exceedingly difficult. Model-fr…

Model-based Reinforcement LearningReinforcement Learning (RL)State Estimation

Learning global control of underactuated systems with Model-Based Reinforcement Learning

2025-04-09 · Niccolò Turcato, Marco Calì, Alberto Dalla Libera, Giulio Giacomuzzo 외

This short paper describes our proposed solution for the third edition of the "AI Olympics with RealAIGym" competition, held at ICRA 2025. We employed Monte-Carlo Probabilistic Inference for Learning Control (MC-PILCO), …

AcrobotModel-based Reinforcement Learning

Probabilistic Pontryagin's Maximum Principle for Continuous-Time Model-Based Reinforcement Learning

2025-04-03 · David Leeftink, Çağatay Yıldız, Steffen Ridderbusch, Max Hinne 외

Without exact knowledge of the true system dynamics, optimal control of non-linear continuous-time systems requires careful treatment of epistemic uncertainty. In this work, we propose a probabilistic extension to Pontry…

Model-based Reinforcement Learningreinforcement-learningReinforcement Learning
1–20 / 708 다음 →