paper-with-me

홈 › Papers

Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection

2025-05-23 · Shrey Pandit, Ashwin Vinod, Liu Leqi, Ying Ding

Aligning large language models (LLMs) to accurately detect hallucinations remains a significant challenge due to the sophisticated nature of hallucinated text. Recognizing that hallucinated samples typically exhibit higher deceptive quality than traditional negative samples, we use these carefully engineered hallucinations as negative examples in the DPO alignment procedure. Our method incorporates a curriculum learning strategy, gradually transitioning the training from easier samples, identified based on the greatest reduction in probability scores from independent fact checking models, to progressively harder ones. This structured difficulty scaling ensures stable and incremental learning. Experimental evaluation demonstrates that our HaluCheck models, trained with curriculum DPO approach and high quality negative samples, significantly improves model performance across various metrics, achieving improvements of upto 24% on difficult benchmarks like MedHallu and HaluEval. Additionally, HaluCheck models demonstrate robustness in zero-shot settings, significantly outperforming larger state-of-the-art models across various benchmarks.

📄 PDF Abstract BibTeX arXiv:2505.17558

Code (0)

등록된 구현이 없습니다.

Tasks

Fact CheckingHallucinationIncremental Learning

Methods 이 논문이 사용한 방법론

DPO 설명 없음

Similar Papers 제목 키워드 기반

Domain and Range Aware Synthetic Negatives Generation for Knowledge Graph Embedding Models

2024-11-22 · Alberto Bernardi, Luca Costabello

Knowledge Graph Embedding models, representing entities and edges in a low-dimensional space, have been extremely successful at solving tasks related to completing and exploring Knowledge Graphs (KGs). One of the key asp…

Graph EmbeddingKnowledge Graph EmbeddingKnowledge Graphs

Curriculum Design for Teaching via Demonstrations: Theory and Applications

2021-06-08 · NeurIPS 2021 12 · Gaurav Yengera, Rati Devidze, Parameswaran Kamalaruban, Adish Singla

We consider the problem of teaching via demonstrations in sequential decision-making settings. In particular, we study how to design a personalized curriculum over demonstrations to speed up the learner's convergence. We…

Decision MakingReinforcement Learning (RL)Sequential Decision Making

How Do Humans Teach: On Curriculum Learning and Teaching Dimension

2011-12-01 · NeurIPS 2011 12 · Faisal Khan, Bilge Mutlu, Jerry Zhu

We study the empirical strategies that humans follow as they teach a target concept with a simple 1D threshold to a robot. Previous studies of computational teaching, particularly the teaching dimension model and the cu…

Teaching Language Models to Hallucinate Less with Synthetic Tasks

2023-10-10 · Erik Jones, Hamid Palangi, Clarisse Simões, Varun Chandrasekaran 외

Large language models (LLMs) frequently hallucinate on abstractive summarization tasks such as document-based question-answering, meeting summarization, and clinical report generation, even though all necessary informati…

Abstractive Text SummarizationHallucinationMeeting SummarizationQuestion Answering+1

Synthetic Hallucinations, Real Gains: Hard Negatives from Frontier Models for FIM Hallucination Mitigation

2026-06-02 · Mahdi Erfanian, Nelson Daniel Troncoso, Aashna Garg, Amabel Gale 외 arxiv

Small open-source code models that power IDE autocomplete still emit hallucinated Fill-in-the-Middle (FIM) completions: syntactically natural calls to methods, parameters, variables, and imports that do not exist in the …