paper-with-me

홈 › Papers

A Novel Automated Curriculum Strategy to Solve Hard Sokoban Planning Instances

2021-10-03 · NeurIPS 2020 12 · Dieqiao Feng, Carla P. Gomes, Bart Selman

In recent years, we have witnessed tremendous progress in deep reinforcement learning (RL) for tasks such as Go, Chess, video games, and robot control. Nevertheless, other combinatorial domains, such as AI planning, still pose considerable challenges for RL approaches. The key difficulty in those domains is that a positive reward signal becomes {\em exponentially rare} as the minimal solution length increases. So, an RL approach loses its training signal. There has been promising recent progress by using a curriculum-driven learning approach that is designed to solve a single hard instance. We present a novel {\em automated} curriculum approach that dynamically selects from a pool of unlabeled training instances of varying task complexity guided by our {\em difficulty quantum momentum} strategy. We show how the smoothness of the task hardness impacts the final learning results. In particular, as the size of the instance pool increases, the ``hardness gap'' decreases, which facilitates a smoother automated curriculum based learning process. Our automated curriculum approach dramatically improves upon the previous approaches. We show our results on Sokoban, which is a traditional PSPACE-complete planning problem and presents a great challenge even for specialized solvers. Our RL agent can solve hard instances that are far out of reach for any previous state-of-the-art Sokoban solver. In particular, our approach can uncover plans that require hundreds of steps, while the best previous search methods would take many years of computing time to solve such instances. In addition, we show that we can further boost the RL performance with an intricate coupling of our automated curriculum approach with a curiosity-driven search strategy and a graph neural net representation.

📄 PDF Abstract BibTeX arXiv:2110.00898

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement LearningReinforcement Learning (RL)Sokoban

Similar Papers 제목 키워드 기반

Solving Hard AI Planning Instances Using Curriculum-Driven Deep Reinforcement Learning

2020-06-04 · Dieqiao Feng, Carla P. Gomes, Bart Selman

Despite significant progress in general AI planning, certain domains remain out of reach of current AI planning systems. Sokoban is a PSPACE-complete planning task and represents one of the hardest domains for current AI…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Sokoban

Procedural Generation of Initial States of Sokoban

2019-07-04 · Dâmaris S. Bento, André G. Pereira, Levi H. S. Lelis

Procedural generation of initial states of state-space search problems have applications in human and machine learning as well as in the evaluation of planning systems. In this paper we deal with the task of generating h…

Sokoban

Transfer Learning and Curriculum Learning in Sokoban

2021-05-25 · Zhao Yang, Mike Preuss, Aske Plaat

Transfer learning can speed up training in machine learning and is regularly used in classification tasks. It reuses prior knowledge from other tasks to pre-train networks for new tasks. In reinforcement learning, learni…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Sokoban+1

Left Heavy Tails and the Effectiveness of the Policy and Value Networks in DNN-based best-first search for Sokoban Planning

2022-06-28 · Dieqiao Feng, Carla Gomes, Bart Selman

Despite the success of practical solvers in various NP-complete domains such as SAT and CSP as well as using deep reinforcement learning to tackle two-player games such as Go, certain classes of PSPACE-hard planning prob…

Deep Reinforcement LearningSokoban

AI in Game Playing: Sokoban Solver

2018-06-29 · Anand Venkatesan, Atishay Jain, Rakesh Grewal

Artificial Intelligence is becoming instrumental in a variety of applications. Games serve as a good breeding ground for trying and testing these algorithms in a sandbox with simpler constraints in comparison to real lif…

AI AgentSokoban