paper-with-me

홈 › Papers

Robust Planning with Compound LLM Architectures: An LLM-Modulo Approach

2024-11-20 · Atharva Gundawar, Karthik Valmeekam, Mudit Verma, Subbarao Kambhampati

Previous work has attempted to boost Large Language Model (LLM) performance on planning and scheduling tasks through a variety of prompt engineering techniques. While these methods can work within the distributions tested, they are neither robust nor predictable. This limitation can be addressed through compound LLM architectures where LLMs work in conjunction with other components to ensure reliability. In this paper, we present a technical evaluation of a compound LLM architecture--the LLM-Modulo framework. In this framework, an LLM is paired with a complete set of sound verifiers that validate its output, re-prompting it if it fails. This approach ensures that the system can never output any fallacious output, and therefore that every output generated is guaranteed correct--something previous techniques have not been able to claim. Our results, evaluated across four scheduling domains, demonstrate significant performance gains with the LLM-Modulo framework using various models. Additionally, we explore modifications to the base configuration of the framework and assess their impact on overall system performance.

📄 PDF Abstract BibTeX arXiv:2411.14484

Code (1)

Atharva-Gundawar/LLM-Modulo-prompts 공식 구현

Tasks

Language ModelingLanguage ModellingLarge Language ModelPrompt EngineeringScheduling

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
BASE 설명 없음

Similar Papers 제목 키워드 기반

Robust Planning with LLM-Modulo Framework: Case Study in Travel Planning

2024-05-31 · Atharva Gundawar, Mudit Verma, Lin Guan, Karthik Valmeekam 외

As the applicability of Large Language Models (LLMs) extends beyond traditional text processing tasks, there is a burgeoning interest in their potential to excel in planning and reasoning assignments, realms traditionall…

On the Planning Abilities of Large Language Models : A Critical Investigation

2023-05-25 · Karthik Valmeekam, Matthew Marquez, Sarath Sreedharan, Subbarao Kambhampati

Intrigued by the claims of emergent reasoning capabilities in LLMs trained on general web corpora, in this paper, we set out to investigate their planning capabilities. We aim to evaluate (1) the effectiveness of LLMs in…

LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks

2024-02-02 · Subbarao Kambhampati, Karthik Valmeekam, Lin Guan, Mudit Verma 외

There is considerable confusion about the role of Large Language Models (LLMs) in planning and reasoning tasks. On one side are over-optimistic claims that LLMs can indeed do these tasks with just the right prompting or …

Modulo Video Recovery via Selective Spatiotemporal Vision Transformer

2025-11-09 · Tianyu Geng, Feng Ji, Wee Peng Tay arxiv

Conventional image sensors have limited dynamic range, causing saturation in high-dynamic-range (HDR) scenes. Modulo cameras address this by folding incident irradiance into a bounded range, yet require specialized unwra…

Video Reconstruction

Linear Temporal Logic Modulo Theories over Finite Traces (Extended Version)

2022-04-28 · Luca Geatti, Alessandro Gianola, Nicola Gigante

This paper studies Linear Temporal Logic over Finite Traces (LTLf) where proposition letters are replaced with first-order formulas interpreted over arbitrary theories, in the spirit of Satisfiability Modulo Theories. Th…