paper-with-me

홈 › Papers

How Far Are LLMs from Symbolic Planners? An NLP-Based Perspective

2025-08-02 · Ma'ayan Armony, Albert Meroño-Peñuela, Gerard Canal arxiv

The reasoning and planning abilities of Large Language Models (LLMs) have been a frequent topic of discussion in recent years. Their ability to take unstructured planning problems as input has made LLMs' integration into AI planning an area of interest. Nevertheless, LLMs are still not reliable as planners, with the generated plans often containing mistaken or hallucinated actions. Existing benchmarking and evaluation methods investigate planning with LLMs, focusing primarily on success rate as a quality indicator in various planning tasks, such as validating plans or planning in relaxed conditions. In this paper, we approach planning with LLMs as a natural language processing (NLP) task, given that LLMs are NLP models themselves. We propose a recovery pipeline consisting of an NLP-based evaluation of the generated plans, along with three stages to recover the plans through NLP manipulation of the LLM-generated plans, and eventually complete the plan using a symbolic planner. This pipeline provides a holistic analysis of LLM capabilities in the context of AI task planning, enabling a broader understanding of the quality of invalid plans. Our findings reveal no clear evidence of underlying reasoning during plan generation, and that a pipeline comprising an NLP-based analysis of the plans, followed by a recovery mechanism, still falls short of the quality and reliability of classical planners. On average, only the first 2.65 actions of the plan are executable, with the average length of symbolically generated plans being 8.4 actions. The pipeline still improves action quality and increases the overall success rate from 21.9% to 27.5%.

📄 PDF Abstract BibTeX arXiv:2508.01300

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Code-as-Symbolic-Planner: Foundation Model-Based Robot Planning via Symbolic Code Generation

2025-03-03 · Yongchao Chen, Yilun Hao, Yang Zhang, Chuchu Fan

Recent works have shown great potentials of Large Language Models (LLMs) in robot task and motion planning (TAMP). Current LLM approaches generate text- or code-based reasoning chains with sub-goals and action plans. How…

Code GenerationCommon Sense ReasoningMotion PlanningTask and Motion Planning

Language Models can Infer Action Semantics for Symbolic Planners from Environment Feedback

2024-06-04 · Wang Zhu, Ishika Singh, Robin Jia, Jesse Thomason

Symbolic planners can discover a sequence of actions from initial to goal states given expert-defined, domain-specific logical action semantics. Large Language Models (LLMs) can directly generate such sequences, but limi…

Vision-Language Interpreter for Robot Task Planning

2023-11-02 · Keisuke Shirai, Cristian C. Beltran-Hernandez, Masashi Hamaya, Atsushi Hashimoto 외

Large language models (LLMs) are accelerating the development of language-guided robot planners. Meanwhile, symbolic planners offer the advantage of interpretability. This paper proposes a new task that bridges these two…

Robot Task PlanningTask Planningvalid

NSP: A Neuro-Symbolic Natural Language Navigational Planner

2024-09-10 · William English, Dominic Simon, Sumit Jha, Rickard Ewetz

Path planners that can interpret free-form natural language instructions hold promise to automate a wide range of robotics applications. These planners simplify user interactions and enable intuitive control over complex…

valid

Hierarchical Planning for Complex Tasks with Knowledge Graph-RAG and Symbolic Verification

2025-04-06 · Cristina Cornelio, Flavio Petruzzellis, Pietro Lio

Large Language Models (LLMs) have shown promise as robotic planners but often struggle with long-horizon and complex tasks, especially in specialized environments requiring external knowledge. While hierarchical planning…

RAGRetrieval-augmented Generation