paper-with-me

홈 › Papers

Optimizing LVLMs with On-Policy Data for Effective Hallucination Mitigation

2025-11-30 · Chengzhi Yu, Yifan Xu, Yifan Chen, Wenyi Zhang arxiv

Recently, large vision-language models (LVLMs) have risen to be a promising approach for multimodal tasks. However, principled hallucination mitigation remains a critical challenge.In this work, we first analyze the data generation process in LVLM hallucination mitigation and affirm that on-policy data significantly outperforms off-policy data, which thus calls for efficient and reliable preference annotation of on-policy data. We then point out that, existing annotation methods introduce additional hallucination in training samples, which may enhance the model's hallucination patterns, to address this problem, we propose training a hallucination classifier giving binary annotations, which guarantee clean chosen samples for the subsequent alignment. To further harness of the power of on-policy data, we design a robust iterative direct preference optimization (DPO) algorithm adopting a dynamic sample reweighting scheme. We conduct comprehensive experiments on three benchmarks with comparison to 8 state-of-the-art baselines. In particular, our approach reduces the hallucination rate of LLaVA-1.5-7B on MMHalBench by 50.8% and the average hallucination rate on Object HalBench by 79.5%; more significantly, our method fully taps into the potential of open-source models, enabling LLaVA-1.5-13B to even surpass the performance of GPT-4V.

📄 PDF Abstract BibTeX arXiv:2512.00706

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Survey of Hallucination in Large Visual Language Models

2024-10-20 · Wei Lan, WenYi Chen, Qingfeng Chen, Shirui Pan 외

The Large Visual Language Models (LVLMs) enhances user interaction and enriches user experience by integrating visual modality on the basis of the Large Language Models (LLMs). It has demonstrated their powerful informat…

HallucinationHallucination EvaluationSurvey

Unified Triplet-Level Hallucination Evaluation for Large Vision-Language Models

2024-10-30 · Junjie Wu, Tsz Ting Chung, Kai Chen, Dit-yan Yeung

Despite the outstanding performance in vision-language reasoning, Large Vision-Language Models (LVLMs) might generate hallucinated contents that do not exist in the given image. Most existing LVLM hallucination benchmark…

HallucinationHallucination EvaluationObjectObject Hallucination+2

R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs

2026-04-22 · Jiahao Xie, Alessio Tonioni, Nathalie Rauschmayr, Federico Tombari 외 arxiv

Large vision-language models (LVLMs) have demonstrated impressive performance in various multimodal understanding and reasoning tasks. However, they still struggle with object hallucinations, i.e., the claim of nonexiste…

Response Generation

Reference-free Hallucination Detection for Large Vision-Language Models

2024-08-11 · Qing Li, Jiahui Geng, Chenyang Lyu, Derui Zhu 외

Large vision-language models (LVLMs) have made significant progress in recent years. While LVLMs exhibit excellent ability in language understanding, question answering, and conversations of visual inputs, they are prone…

HallucinationQuestion AnsweringUncertainty Quantification

Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding

2025-02-03 · Chao Wang, Xuancheng Zhou, Weiwei Fu, Yang Zhou

Large Visual Language Models (LVLMs) integrate visual and linguistic modalities, exhibiting exceptional performance across various multimodal tasks. Nevertheless, LVLMs remain vulnerable to the issue of object hallucinat…

AttributeMMEObject