paper-with-me

홈 › Papers

Teaching Language Models to Hallucinate Less with Synthetic Tasks

2023-10-10 · Erik Jones, Hamid Palangi, Clarisse Simões, Varun Chandrasekaran, Subhabrata Mukherjee, Arindam Mitra, Ahmed Awadallah, Ece Kamar

Large language models (LLMs) frequently hallucinate on abstractive summarization tasks such as document-based question-answering, meeting summarization, and clinical report generation, even though all necessary information is included in context. However, optimizing LLMs to hallucinate less on these tasks is challenging, as hallucination is hard to efficiently evaluate at each optimization step. In this work, we show that reducing hallucination on a synthetic task can also reduce hallucination on real-world downstream tasks. Our method, SynTra, first designs a synthetic task where hallucinations are easy to elicit and measure. It next optimizes the LLM's system message via prefix-tuning on the synthetic task, and finally transfers the system message to realistic, hard-to-optimize tasks. Across three realistic abstractive summarization tasks, SynTra reduces hallucination for two 13B-parameter LLMs using only a synthetic retrieval task for supervision. We also find that optimizing the system message rather than the model weights can be critical; fine-tuning the entire model on the synthetic task can counterintuitively increase hallucination. Overall, SynTra demonstrates that the extra flexibility of working with synthetic data can help mitigate undesired behaviors in practice.

📄 PDF Abstract BibTeX arXiv:2310.06827

Code (0)

등록된 구현이 없습니다.

Tasks

Abstractive Text SummarizationHallucinationMeeting SummarizationQuestion AnsweringRetrieval

Similar Papers 제목 키워드 기반

Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection

2025-05-23 · Shrey Pandit, Ashwin Vinod, Liu Leqi, Ying Ding

Aligning large language models (LLMs) to accurately detect hallucinations remains a significant challenge due to the sophisticated nature of hallucinated text. Recognizing that hallucinated samples typically exhibit high…

Fact CheckingHallucinationIncremental Learning

VLMs Need Words: Vision Language Models Ignore Visual Detail In Favor of Semantic Anchors

2026-04-02 · Haz Sameen Shahgir, Xiaofu Chen, Yu Fu, Erfan Shayegani 외 arxiv

Vision-language models (VLMs) have achieved impressive performance across a wide range of multimodal tasks. However, they often fail on tasks that require fine-grained visual perception, even when the required informatio…

Semantic correspondenceMultimodal Reasoning

Teaching Humans When To Defer to a Classifier via Exemplars

2021-11-22 · Hussein Mozannar, Arvind Satyanarayan, David Sontag

Expert decision makers are starting to rely on data-driven automated agents to assist them with various tasks. For this collaboration to perform properly, the human decision maker must have a mental model of when and whe…

Multi-hop Question AnsweringQuestion Answeringvalid

MedVeriSeg: Teaching LISA-Like Medical Segmentation Models to Verify Query Validity Without Extra Training

2026-04-11 · Qinyue Tong, Xiaozhen Wang, Ziqian Lu, Jun Liu 외 arxiv

Despite recent progress in text-prompt-based medical image segmentation, existing LISA-like MLLM-based methods typically generate masks regardless of whether the target specified in the query is present, leading to hallu…

Medical Image Segmentation

The Online Pivot: Lessons Learned from Teaching a Text and Data Mining Course in Lockdown, Enhancing online Teaching with Pair Programming and Digital Badges

2021-05-03 · NAACL (TeachingNLP) 2021 6 · Beatrice Alex, Clare Llewellyn, Pawel Michal Orzechowski, Maria Boutchkova

In this paper we provide an account of how we ported a text and data mining course online in summer 2020 as a result of the COVID-19 pandemic and how we improved it in a second pilot run. We describe the course, how we a…