paper-with-me

홈 › Papers

Planning and Editing What You Retrieve for Enhanced Tool Learning

2024-03-30 · Tenghao Huang, Dongwon Jung, Muhao Chen

Recent advancements in integrating external tools with Large Language Models (LLMs) have opened new frontiers, with applications in mathematical reasoning, code generators, and smart assistants. However, existing methods, relying on simple one-time retrieval strategies, fall short on effectively and accurately shortlisting relevant tools. This paper introduces a novel PLUTO (Planning, Learning, and Understanding for TOols) approach, encompassing Plan-and-Retrieve (P&R) and Edit-and-Ground (E&G) paradigms. The P&R paradigm consists of a neural retrieval module for shortlisting relevant tools and an LLM-based query planner that decomposes complex queries into actionable tasks, enhancing the effectiveness of tool utilization. The E&G paradigm utilizes LLMs to enrich tool descriptions based on user scenarios, bridging the gap between user queries and tool functionalities. Experiment results demonstrate that these paradigms significantly improve the recall and NDCG in tool retrieval tasks, significantly surpassing current state-of-the-art models.

📄 PDF Abstract BibTeX arXiv:2404.00450

Code (1)

tenghaohuang/pluto 공식 구현

Tasks

Mathematical ReasoningRetrieval

Similar Papers 제목 키워드 기반

Aurora: Unified Video Editing with a Tool-Using Agent

2026-05-18 · Yongsheng Yu, Ziyun Zeng, Zhiyuan Xiao, Zhenghong Zhou 외 arxiv

Recent video editing models have converged on a unified conditioning design: a single diffusion transformer jointly consumes text, source video, and reference images, and one set of weights covers replacement, removal, s…

Style Transfer

Agentic Planning with Reasoning for Image Styling via Offline RL

2026-03-07 · Subhojyoti Mukherjee, Stefano Petrangeli, Branislav Kveton, Trung Bui 외 arxiv

Direct prompt-based editing often fails on complex transformations because vague and subjective prompts often require nuanced understanding of what should be changed in the image. Our core intuition is that leveraging co…

Synthetic Data GenerationInstruction FollowingImage EditingOffline RL

ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL

2026-08-28 · Zhuoshi Pan, Qizhi Pei, Junru Lu, Honglin Lin 외 arxiv

Long-horizon agentic tasks require large language models (LLMs) to iteratively retrieve, integrate, and maintain dispersed information across multi-turn interactions, but preserving all interaction histories leads to a c…

Cross-lingual neural fuzzy matching for exploiting target-language monolingual corpora in computer-aided translation

2024-01-16 · Miquel Esplà-Gomis, Víctor M. Sánchez-Cartagena, Juan Antonio Pérez-Ortiz, Felipe Sánchez-Martínez

Computer-aided translation (CAT) tools based on translation memories (MT) play a prominent role in the translation workflow of professional translators. However, the reduced availability of in-domain TMs, as compared to …

SentenceSentence EmbeddingsTranslation

Toolshed: Scale Tool-Equipped Agents with Advanced RAG-Tool Fusion and Tool Knowledge Bases

2024-10-18 · Elias Lumer, Vamse Kumar Subbiah, James A. Burke, Pradeep Honaganahalli Basavaraju 외

Recent advancements in tool-equipped Agents (LLMs) have enabled complex tasks like secure database interactions and multi-agent code development. However, scaling tool capacity beyond agent reasoning or model limits rema…

RAGRetrievalRetrieval-augmented Generation