paper-with-me

홈 › Papers

Train for Truth, Keep the Skills: Binary Retrieval-Augmented Reward Mitigates Hallucinations

2025-10-20 · Tong Chen, Akari Asai, Luke Zettlemoyer, Hannaneh Hajishirzi, Faeze Brahman arxiv

Language models often generate factually incorrect information unsupported by their training data, a phenomenon known as extrinsic hallucination. Existing mitigation approaches often degrade performance on open-ended generation and downstream tasks, limiting their practical utility. We propose an online reinforcement learning method using a novel binary retrieval-augmented reward (RAR) to address this tradeoff. Unlike continuous reward schemes, our approach assigns a reward of one only when the model's output is entirely factually correct, and zero otherwise. We evaluate our method on Qwen3 reasoning models across diverse tasks. For open-ended generation, binary RAR achieves a 39.3% reduction in hallucination rates, substantially outperforming both supervised training and continuous-reward RL baselines. In short-form question answering, the model learns calibrated abstention, strategically outputting "I don't know" when faced with insufficient parametric knowledge. This yields 44.4% and 21.7% fewer incorrect answers on PopQA and GPQA, respectively. Crucially, these factuality gains come without performance degradation on instruction following, math, or code, whereas continuous-reward RL, despite improving factuality, induces quality regressions.

📄 PDF Abstract BibTeX arXiv:2510.17733

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningInstruction FollowingQuestion Answering

Similar Papers 제목 키워드 기반

Demystifying Agent Skills: Why They Work-Until They Don't

2026-08-14 · Zhiyuan Jiang, Fangrui Huang, Hanwen Xing, Xander Wu 외 arxiv

Skills have emerged as a practical and effective approach for enhancing LLM agents at inference time through structured packages of knowledge. However, existing evaluations largely measure whether skills improve aggregat…

Binary Codes for Tagging X-Ray Images via Deep De-Noising Autoencoders

2016-04-24 · Antonio Sze-To, Hamid. R. Tizhoosh, Andrew K. C. Wong

A Content-Based Image Retrieval (CBIR) system which identifies similar medical images based on a query image can assist clinicians for more accurate diagnosis. The recent CBIR research trend favors the construction and u…

Content-Based Image RetrievalImage RetrievalRetrieval

SkillDAG: Self-Evolving Typed Skill Graphs for LLM Skill Selection at Scale

2026-06-02 · Tong Bai, Zhenglin Wan, Pengfei Zhou, Xingrui Yu 외 arxiv

As LLM agents adopt large skill libraries, selecting the right subset becomes a structural problem rather than a similarity-matching one: skills depend on, conflict with, specialize, or duplicate one another, a structure…

Eye Movement Feature Classification for Soccer Goalkeeper Expertise Identification in Virtual Reality

2020-09-23 · Benedikt Hosp, Florian Schultz, Oliver Höner, Enkelejda Kasneci

The latest research in expertise assessment of soccer players has affirmed the importance of perceptual skills (especially for decision making) by focusing either on high experimental control or on a realistic presentati…

Decision MakingGeneral Classification

DisTop: Discovering a Topological representation to learn diverse and rewarding skills

2021-06-06 · Arthur Aubret, Laetitia Matignon, Salima Hassas

The optimal way for a deep reinforcement learning (DRL) agent to explore is to learn a set of skills that achieves a uniform distribution of states. Following this,we introduce DisTop, a new model that simultaneously lea…

Deep Reinforcement LearningHierarchical Reinforcement LearningMuJoCoreinforcement-learning+3