paper-with-me

홈 › Papers

Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Language Models

2025-05-18 · Kai Tang, Jinhao You, Xiuqi Ge, Hanze Li, Yichen Guo, Xiande Huang

Despite the impressive capabilities of Large Vision-Language Models (LVLMs), they remain susceptible to hallucinations-generating content that is inconsistent with the input image. Existing training-free hallucination mitigation methods often suffer from unstable performance and high sensitivity to hyperparameter settings, limiting their practicality and broader adoption. In this paper, we propose a novel decoding mechanism, Decoding with Inter-layer Consistency via Layer Aggregation (DCLA), which requires no retraining, fine-tuning, or access to external knowledge bases. Specifically, our approach constructs a dynamic semantic reference by aggregating representations from previous layers, and corrects semantically deviated layers to enforce inter-layer consistency. The method allows DCLA to robustly mitigate hallucinations across multiple LVLMs. Experiments on hallucination benchmarks such as MME and POPE demonstrate that DCLA effectively reduces hallucinations while enhancing the reliability and performance of LVLMs.

📄 PDF Abstract BibTeX arXiv:2505.12343

Code (0)

등록된 구현이 없습니다.

Tasks

HallucinationMME

Similar Papers 제목 키워드 기반

MAP: Mitigating Hallucinations in Large Vision-Language Models with Map-Level Attention Processing

2025-08-03 · Chenxi Li, Yichen Guo, Benfang Qian, Jinhao You 외 arxiv

Large Vision-Language Models (LVLMs) have achieved impressive performance in multimodal tasks, but they still suffer from hallucinations, i.e., generating content that is grammatically accurate but inconsistent with visu…

Mitigating Hallucinations in Large Language Models Via Decoder Layer Skipping

2026-05-30 · Hanze Li, Jinhao You, Yichen Guo, Kai Tang 외 arxiv

Large Language Models (LLMs) have achieved strong performance across diverse natural language tasks, yet their outputs often suffer from hallucinations -- content that is misaligned with factual information. In this work…

Listen to the Layers: Mitigating Hallucinations with Inter-Layer Disagreement

2026-02-10 · Koduvayur Subbalakshmi, Sabbir Hossain Ujjal, Venkata Krishna Teja Mangichetty, Nastaran Jamalipour Soofi arxiv

Pretrained Large Language Models (LLMs) are prone to generating fluent yet factually incorrect text-a phenomenon known as hallucinations, undermining their reliability and utility in downstream tasks. We hypothesize that…

Mathematical ReasoningCode Generation

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering

2026-07-05 · Kai Tang, Jinhao You, Bohua Zhang, Yichen Guo 외 arxiv

Large Vision-Language Models (LVLMs) have achieved remarkable progress in visual understanding tasks such as image captioning and visual question answering. However, they remain susceptible to hallucinations, generating …

Visual Question AnsweringFeature EngineeringImage Captioning

MRFD: Multi-Region Fusion Decoding with Self-Consistency for Mitigating Hallucinations in LVLMs

2025-08-14 · Haonan Ge, Yiwei Wang, Ming-Hsuan Yang, Yujun Cai arxiv

Large Vision-Language Models (LVLMs) have shown strong performance across multimodal tasks. However, they often produce hallucinations -- text that is inconsistent with visual input, due to the limited ability to verify …