paper-with-me

홈 › Papers

Goal Reasoning by Selecting Subgoals with Deep Q-Learning

2020-12-22 · Carlos Núñez-Molina, Vladislav Nikolov, Ignacio Vellido, Juan Fernández-Olivares

In this work we propose a goal reasoning method which learns to select subgoals with Deep Q-Learning in order to decrease the load of a planner when faced with scenarios with tight time restrictions, such as online execution systems. We have designed a CNN-based goal selection module and trained it on a standard video game environment, testing it on different games (planning domains) and levels (planning problems) to measure its generalization abilities. When comparing its performance with a satisfying planner, the results obtained show both approaches are able to find plans of good quality, but our method greatly decreases planning time. We conclude our approach can be successfully applied to different types of domains (games), and shows good generalization properties when evaluated on new levels (problems) of the same game (domain).

📄 PDF Abstract BibTeX arXiv:2012.12335

Code (0)

등록된 구현이 없습니다.

Tasks

Q-Learning

Methods 이 논문이 사용한 방법론

Q-Learning Q-Learning is an off-policy temporal difference control algorithm: $$Q\left(S\_{t}, A\_{t}\right) \leftarrow Q\left(S\_{t}, A\_{t}\right) + \alpha\left[R_{t+1} +…

Similar Papers 제목 키워드 기반

Visual scoping operations for physical assembly

2021-06-10 · Felix J Binder, Marcelo M Mattar, David Kirsh, Judith E Fan

Planning is hard. The use of subgoals can make planning more tractable, but selecting these subgoals is computationally costly. What algorithms might enable us to reap the benefits of planning using subgoals while minimi…

Fast and Precise: Adjusting Planning Horizon with Adaptive Subgoal Search

2022-06-01 · Michał Zawalski, Michał Tyrolski, Konrad Czechowski, Tomasz Odrzygóźdź 외

Complex reasoning problems contain states that vary in the computational cost required to determine a good action plan. Taking advantage of this property, we propose Adaptive Subgoal Search (AdaSubS), a search method tha…

Rubik's CubeSokoban

Efficient Navigation in Unknown Indoor Environments with Vision-Language Models

2025-10-06 · D. Schwartz, K. Kondo, J. P. How arxiv

We present a novel high-level planning framework that leverages vision-language models (VLMs) to improve autonomous navigation in unknown indoor environments with many dead ends. Traditional exploration methods often tak…

Subgoal Search For Complex Reasoning Tasks

2021-08-25 · NeurIPS 2021 12 · Konrad Czechowski, Tomasz Odrzygóźdź, Marek Zbysiński, Michał Zawalski 외

Humans excel in solving complex reasoning tasks through a mental process of moving from one idea to a related one. Inspired by this, we propose Subgoal Search (kSubS) method. Its key component is a learned subgoal genera…

DiversityRubik's CubeSokoban

Goal-Conditioned Reinforcement Learning with Imagined Subgoals

2021-07-01 · Elliot Chane-Sane, Cordelia Schmid, Ivan Laptev

Goal-conditioned reinforcement learning endows an agent with a large variety of skills, but it often struggles to solve tasks that require more temporally extended reasoning. In this work, we propose to incorporate imagi…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)