paper-with-me

홈 › Papers

Selective Dyna-style Planning Under Limited Model Capacity

2020-07-05 · ICML 2020 1 · Zaheer Abbas, Samuel Sokota, Erin J. Talvitie, Martha White

In model-based reinforcement learning, planning with an imperfect model of the environment has the potential to harm learning progress. But even when a model is imperfect, it may still contain information that is useful for planning. In this paper, we investigate the idea of using an imperfect model selectively. The agent should plan in parts of the state space where the model would be helpful but refrain from using the model where it would be harmful. An effective selective planning mechanism requires estimating predictive uncertainty, which arises out of aleatoric uncertainty, parameter uncertainty, and model inadequacy, among other sources. Prior work has focused on parameter uncertainty for selective planning. In this work, we emphasize the importance of model inadequacy. We show that heteroscedastic regression can signal predictive uncertainty arising from model inadequacy that is complementary to that which is detected by methods designed for parameter uncertainty, indicating that considering both parameter uncertainty and model inadequacy may be a more promising direction for effective selective planning than either in isolation.

📄 PDF Abstract BibTeX arXiv:2007.02418

Code (0)

등록된 구현이 없습니다.

Tasks

modelModel-based Reinforcement Learning

Similar Papers 제목 키워드 기반

MEPG:Multi-Expert Planning and Generation for Compositionally-Rich Image Generation

2025-09-04 · Yuan Zhao, Lin Liu arxiv

Text-to-image diffusion models have achieved remarkable image quality, but they still struggle with complex, multiele ment prompts, and limited stylistic diversity. To address these limitations, we propose a Multi-Expert…

Image Generation

Leveraging the Value of Information in POMDP Planning

2026-04-01 · Zakariya Laouar, Qi Heng Ho, Zachary Sunberg arxiv

Partially observable Markov decision processes (POMDPs) offer a principled formalism for planning under state and transition uncertainty. Despite advances made towards solving large POMDPs, obtaining performant policies …

Dyna-Style Planning with Linear Function Approximation and Prioritized Sweeping

2012-06-13 · Richard S. Sutton, Csaba Szepesvari, Alborz Geramifard, Michael P. Bowling

We consider the problem of efficiently learning optimal control policies and value functions over large state spaces in an online setting in which estimates must be available after each interaction with the world. This p…

PLAN-S: Bridging Planning with Latent Style Dynamics for Autonomous Driving World Models

2026-06-04 · Xiaoyun Qiu, Jingtao He, Yijie Chen, Yusong Huang 외 arxiv

Latent world models (LWMs) have strengthened end-to-end autonomous driving by forecasting compact scene dynamics for downstream planning. However, existing LWM-based planners usually generate trajectories directly from e…

Autonomous Driving

Self-Steering Deep Non-Linear Spatially Selective Filters for Efficient Extraction of Moving Speakers under Weak Guidance

2025-07-03 · Jakob Kienegger, Alina Mannanova, Huajian Fang, Timo Gerkmann arxiv

Recent works on deep non-linear spatially selective filters demonstrate exceptional enhancement performance with computationally lightweight architectures for stationary speakers of known directions. However, to maintain…

Speech Enhancement