paper-with-me

홈 › Papers

DSL-R1: From SQL to DSL for Training Retrieval Agents across Structured and Unstructured Data with Reinforcement Learning

2026-01-14 · Yunhai Hu, Junwei Zhou, Yumo Cao, Yitao Long, Yiwei Xu, Qiyi Jiang, Weiyao Wang, Xiaoyu Cao, Zhen Sun, Yiran Zou, Nan Du arxiv

Effective retrieval in complex domains requires bridging the gap between structured metadata and unstructured content. Existing systems typically isolate these capabilities, relying on either symbolic filtering or vector similarity, failing to capture their interplay. In this work, we propose DSL-R1, a unified framework that synergizes logical reasoning with semantic matching via a novel Domain-Specific Language (DSL). By embedding vector primitives within SQL-style operators, our approach leverages the complementary strengths of symbolic precision and semantic coverage. We further introduce a reinforcement learning mechanism where rule-based execution feedback and retrieval quality rewards jointly optimize the DSL generation, balancing structural correctness and semantic alignment. Evaluations on a large-scale industrial email benchmark demonstrate that DSL-R1 achieves a +12.3% improvement in Hit@1/3, consistently outperforming decoupled baselines and establishing a robust paradigm for hybrid retrieval.

📄 PDF Abstract BibTeX arXiv:2603.21018

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningLogical Reasoning

Similar Papers 제목 키워드 기반

$τ$-Knowledge: Evaluating Conversational Agents over Unstructured Knowledge

2026-03-04 · Quan Shi, Alexandra Zytek, Pedram Razavi, Karthik Narasimhan 외 arxiv

Conversational agents are increasingly deployed in knowledge-intensive settings, where correct behavior depends on retrieving and applying domain-specific knowledge from large, proprietary, and unstructured corpora durin…

Dynamic Multi-Agent Orchestration and Retrieval for Multi-Source Question-Answer Systems using Large Language Models

2024-12-23 · Antony Seabra, Claudio Cavalcante, Joao Nepomuceno, Lucas Lago 외

We propose a methodology that combines several advanced techniques in Large Language Model (LLM) retrieval to support the development of robust, multi-source question-answer systems. This methodology is designed to integ…

Language ModelingLanguage ModellingLarge Language ModelPrompt Engineering+3

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

2026-05-27 · Yuting Xu, Jiayi Tian, Jian Liang, Xin Xiong 외 arxiv

Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomous Agents continue to advance, their evaluation must evolve beyond si…

Structuring the Unstructured: A Multi-Agent System for Extracting and Querying Financial KPIs and Guidance

2025-05-25 · Chanyeol Choi, Jihoon Kwon, Minjae Kim, Juneha Hwang 외

Extracting structured and quantitative insights from unstructured financial filings is essential in investment research, yet remains time-consuming and resource-intensive. Conventional approaches in practice rely heavily…

Natural Language QueriesRetrievalText to SQLText-To-SQL

Retrieval-Augmented Machine Translation with Unstructured Knowledge

2024-12-05 · Jiaan Wang, Fandong Meng, Yingxue Zhang, Jie zhou

Retrieval-augmented generation (RAG) introduces additional information to enhance large language models (LLMs). In machine translation (MT), previous work typically retrieves in-context examples from paired MT corpora, o…

Knowledge GraphsMachine TranslationRAGRetrieval+3