paper-with-me

홈 › Papers

A Unifying Framework for Reinforcement Learning and Planning

2020-06-26 · Thomas M. Moerland, Joost Broekens, Aske Plaat, Catholijn M. Jonker

Sequential decision making, commonly formalized as optimization of a Markov Decision Process, is a key challenge in artificial intelligence. Two successful approaches to MDP optimization are reinforcement learning and planning, which both largely have their own research communities. However, if both research fields solve the same problem, then we might be able to disentangle the common factors in their solution approaches. Therefore, this paper presents a unifying algorithmic framework for reinforcement learning and planning (FRAP), which identifies underlying dimensions on which MDP planning and learning algorithms have to decide. At the end of the paper, we compare a variety of well-known planning, model-free and model-based RL algorithms along these dimensions. Altogether, the framework may help provide deeper insight in the algorithmic design space of planning and reinforcement learning.

📄 PDF Abstract BibTeX arXiv:2006.15009

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)Sequential Decision Making

Similar Papers 제목 키워드 기반

Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning

2025-06-23 · Anthony Kobanda, Waris Radji, Mathieu Petitbois, Odalric-Ambrym Maillard 외

Offline Goal-Conditioned Reinforcement Learning seeks to train agents to reach specified goals from previously collected trajectories. Scaling that promises to long-horizon tasks remains challenging, notably due to compo…

Metric Learningreinforcement-learningReinforcement Learning

Toward Greater Autonomy in Materials Discovery Agents: Unifying Planning, Physics, and Scientists

2025-06-05 · Lianhao Zhou, Hongyi Ling, Keqiang Yan, Kaiji Zhao 외

We aim at designing language agents with greater autonomy for crystal materials discovery. While most of existing studies restrict the agents to perform specific tasks within predefined workflows, we aim to automate work…

RulePlanner: All-in-One Reinforcement Learner for Unifying Design Rules in 3D Floorplanning

2026-01-30 · Ruizhe Zhong, Xingbo Du, Junchi Yan arxiv

Floorplanning determines the coordinate and shape of each module in Integrated Circuits. With the scaling of technology nodes, in floorplanning stage especially 3D scenarios with multiple stacked layers, it has become in…

Reinforcement Learning

Model Predictive Adversarial Imitation Learning for Planning from Observation

2025-07-29 · Tyler Han, Yanda Bao, Bhaumik Mehta, Gabriel Guo 외 arxiv

Human demonstration data is often ambiguous and incomplete, motivating imitation learning approaches that also exhibit reliable planning behavior. A common paradigm to perform planning-from-demonstration involves learnin…

Reinforcement Learning

Hierarchical Active Inference using Successor Representations

2026-04-17 · Prashant Rangarajan, Rajesh P. N. Rao arxiv

Active inference, a neurally-inspired model for inferring actions based on the free energy principle (FEP), has been proposed as a unifying framework for understanding perception, action, and learning in the brain. Activ…

Reinforcement Learning