paper-with-me

홈 › Papers

Code-Driven Planning in Grid Worlds with Large Language Models

2025-05-15 · Ashwath Vaithinathan Aravindan, Zhisheng Tang, Mayank Kejriwal

We propose an iterative programmatic planning (IPP) framework for solving grid-based tasks by synthesizing interpretable agent policies expressed in code using large language models (LLMs). Instead of relying on traditional search or reinforcement learning, our approach uses code generation as policy synthesis, where the LLM outputs executable programs that map environment states to action sequences. Our proposed architecture incorporates several prompting strategies, including direct code generation, pseudocode-conditioned refinement, and curriculum-based prompting, but also includes an iterative refinement mechanism that updates code based on task performance feedback. We evaluate our approach using six leading LLMs and two challenging grid-based benchmarks (GRASP and MiniGrid). Our IPP framework demonstrates improvements over direct code generation ranging from 10\% to as much as 10x across five of the six models and establishes a new state-of-the-art result for GRASP. IPP is found to significantly outperform direct elicitation of a solution from GPT-o3-mini (by 63\% on MiniGrid to 116\% on GRASP), demonstrating the viability of the overall approach. Computational costs of all code generation approaches are similar. While code generation has a higher initial prompting cost compared to direct solution elicitation (\$0.08 per task vs. \$0.002 per instance for GPT-o3-mini), the code can be reused for any number of instances, making the amortized cost significantly lower (by 400x on GPT-o3-mini across the complete GRASP benchmark).

📄 PDF Abstract BibTeX arXiv:2505.10749

Code (0)

등록된 구현이 없습니다.

Tasks

Code Generation

Similar Papers 제목 키워드 기반

WorldCoder, a Model-Based LLM Agent: Building World Models by Writing Code and Interacting with the Environment

2024-02-19 · Hao Tang, Darren Key, Kevin Ellis

We give a model-based agent that builds a Python program representing its knowledge of the world based on its interactions with the environment. The world model tries to explain its interactions, while also being optimis…

Program SynthesisTask Planning

REMI: Reconstructing Episodic Memory During Internally Driven Path Planning

2025-07-02 · Zhaoze Wang, Genela Morris, Dori Derdikman, Pratik Chaudhari 외 arxiv

Grid cells in the medial entorhinal cortex (MEC) and place cells in the hippocampus (HC) both form spatial representations. Grid cells fire in triangular grid patterns, while place cells fire at specific locations and re…

OpenSpiel: A Framework for Reinforcement Learning in Games

2019-08-26 · Marc Lanctot, Edward Lockhart, Jean-Baptiste Lespiau, Vinicius Zambaldi 외

OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games. OpenSpiel supports n-player (single- and multi- agent) zero-sum, cooperative and gener…

General Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Graph neural induction of value iteration

2020-09-26 · Andreea Deac, Pierre-Luc Bacon, Jian Tang

Many reinforcement learning tasks can benefit from explicit planning based on an internal model of the environment. Previously, such planning components have been incorporated through a neural network that partially alig…

Deep Reinforcement LearningGraph Neural Networkreinforcement-learningReinforcement Learning+1

Data-driven, metaheuristic-based off-grid microgrid capacity planning optimisation and scenario analysis: Insights from a case study of Aotea-Great Barrier Island

2022-09-21 · Soheil Mohseni, Roomana Khalid, Alan C Brent

Small privately-purchased off-grid renewable energy systems (RESs) are increasingly used for energy generation in remote areas. However, such privately-purchased stand-alone RESs are often unaffordable for households wit…