paper-with-me

홈 › Papers

ToolPlanner: A Tool Augmented LLM for Multi Granularity Instructions with Path Planning and Feedback

2024-09-23 · Qinzhuo Wu, Wei Liu, Jian Luan, Bin Wang

Recently, tool-augmented LLMs have gained increasing attention. Given an instruction, tool-augmented LLMs can interact with various external tools in multiple rounds and provide a final answer. However, previous LLMs were trained on overly detailed instructions, which included API names or parameters, while real users would not explicitly mention these API details. This leads to a gap between trained LLMs and real-world scenarios. In addition, most works ignore whether the interaction process follows the instruction. To address these issues, we constructed a training dataset called MGToolBench, which contains statement and category-level instructions to better reflect real-world scenarios. In addition, we propose ToolPlanner, a two-stage reinforcement learning framework that utilizes path planning and two feedback mechanisms to enhance the LLM's task completion and instruction-following capabilities. Experimental results show that ToolPlanner significantly improves the Match Rate, Pass Rate and Win Rate by 26.8%, 20.2%, and 5.6% compared to the SOTA model. Human evaluation verifies that the multi-granularity instructions can better align with users' usage habits. Our data and code will be released upon acceptance.

📄 PDF Abstract BibTeX arXiv:2409.14826

Code (1)

xiaomi/toolplanner 공식 구현 pytorch

Tasks

Instruction Following

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Empowering Large Language Models for Textual Data Augmentation

2024-04-26 · Yichuan Li, Kaize Ding, Jianling Wang, Kyumin Lee

With the capabilities of understanding and executing natural language instructions, Large language models (LLMs) can potentially act as a powerful tool for textual data augmentation. However, the quality of augmented dat…

Data AugmentationDiversityFew-Shot Learning

OSGuard: A Benchmark for Safety in Computer-Use Agents

2026-06-13 · Mina Mohammadmirzaei, Jeffrey Flanigan arxiv

Computer-use agents are increasingly evaluated by whether they complete realistic desktop and web tasks. However, task success alone can miss failures in which an agent reaches the nominal goal through an unsafe shortcut…

HyperTool: Beyond Step-Wise Tool Calls for Tool-Augmented Agents

2026-06-11 · Yaxin Du, Yifan Zhou, Yujie Ge, Jiajun Wang 외 arxiv

Tool-augmented LLM agents commonly rely on step-wise atomic tool calls, where each invocation, observation, and value transfer is exposed in the main reasoning trace. This creates an \emph{execution-granularity mismatch}…

Language-based Photo Color Adjustment for Graphic Designs

2023-08-06 · Zhenwei Wang, Nanxuan Zhao, Gerhard Hancke, Rynson W. H. Lau

Adjusting the photo color to associate with some design elements is an essential way for a graphic design to effectively deliver its message and make it aesthetically pleasing. However, existing tools and previous works …

JTPRO: A Joint Tool-Prompt Reflective Optimization Framework for Language Agents

2026-04-20 · Sandip Ghoshal, Anshul Mittal, Jyotika Singh, Miguel Ballesteros 외 arxiv

Large language model (LLM) agents augmented with external tools often struggle as number of tools grow large and become domain-specific. In such settings, ambiguous tool descriptions and under-specified agent instruction…

Slot Filling