paper-with-me

홈 › Papers

From Evidence to Trajectory: Abductive Reasoning Path Synthesis for Training Retrieval-Augmented Generation Agents

2025-09-27 · Muzhi Li, Jinhu Qi, Yihong Wu, Minghao Zhao, Liheng Ma, Yifan Li, Xinyu Wang, Yingxue Zhang, Ho-fung Leung, Irwin King arxiv

Retrieval-augmented generation agents development is hindered by the lack of process-level supervision to effectively guide agentic capabilities like task decomposition, retriever invocation, and stepwise decision-making. While reinforcement learning offers a potential solution, it suffers from sparse rewards and the limited reasoning capabilities of large language models (LLMs). Meanwhile, existing data synthesis methods only produce chain-of-thought rationales and fail to model environmental interactions. In this paper, we propose EviPath, an evidence-anchored reasoning path synthesis paradigm for RAG agent development. EviPath comprises: (i) Abductive Subtask Planning, which decomposes the problem into sub-questions and iteratively plans an optimal solution path based on the dependencies between them; (ii) Faithful Sub-question Answering, which uses supporting evidence to construct a proxy environment to generate reasoning thoughts and answers for each sub-question; and (iii) Conversational Fine-Tuning, which formats the complete agent-environment interaction trajectory into a dialogue format suitable for Supervised Fine-Tuning. EviPath allows LLMs to learn complex reasoning and tool-use capabilities directly from synthesized data. Extensive experiments on widely-used question-answering benchmarks show that an 8B parameter model trained with EviPath-synthesized data significantly and consistently outperforms state-of-the-art baselines with a double-digit absolute EM gain of 14.7% in open-domain question answering.

📄 PDF Abstract BibTeX arXiv:2509.23071

Code (0)

등록된 구현이 없습니다.

Tasks

Open-Domain Question AnsweringReinforcement Learning

Similar Papers 제목 키워드 기반

Abductive Inference in Retrieval-Augmented Language Models: Generating and Validating Missing Premises

2025-11-06 · Shiyin Lin arxiv

Large Language Models (LLMs) enhanced with retrieval -- commonly referred to as Retrieval-Augmented Generation (RAG) -- have demonstrated strong performance in knowledge-intensive tasks. However, RAG pipelines often fail…

Neurosymbolic Clinical Trial Matching via LLM-Driven Abduction and Logical Verification

2026-06-18 · Baiyang Qu, Leonardo Ranaldi, Xi Wang, Marco Valentino arxiv

Large Language Models (LLMs) offer a promising path to automate Clinical Trial Matching (CTM), but still struggle with the deterministic verification required for complex eligibility criteria. Conversely, purely symbolic…

SemEval-2026 Task 12: Abductive Event Reasoning: Towards Real-World Event Causal Inference for Large Language Models

2026-03-23 · Pengfei Cao, Mingxuan Yang, Yubo Chen, Chenlong Zhang 외 arxiv

Understanding why real-world events occur is important for both natural language processing and practical decision-making, yet direct-cause inference remains underexplored in evidence-rich settings. To address this gap, …

Causal Inference

Don't Let Me Ask for It: LLMs Show Deficiencies in Active Multi-Turn Information Acquisition for Abductive Inference

2026-08-04 · Shahrukh Mohiuddin, Chalamalasetti Kranti, Sherzod Hakimov, David Schlangen arxiv

Abductive reasoning requires forming hypotheses that explain observed evidence and revising them as new evidence becomes available. While large language models (LLMs) are often evaluated on whether they solve abductive r…

Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs

2026-04-09 · Moein Salimi, Shaygan Adim, Danial Parnian, Nima Alighardashi 외 arxiv

Regardless of its foundational role in human discovery and sense-making, abductive reasoning--the inference of the most plausible explanation for an observation--has been relatively underexplored in Large Language Models…