paper-with-me

홈 › Papers

Contrastive Decoding Mitigates Score Range Bias in LLM-as-a-Judge

2025-10-21 · Yoshinari Fujinuma arxiv

Large Language Models (LLMs) are commonly used as evaluators in various applications, but the reliability of the outcomes remains a challenge. One such challenge is using LLMs-as-judges for direct assessment, i.e., assigning scores from a specified range without any references. Focusing on summarization, we first show that this challenge stems from LLM judge outputs being associated with score range bias, i.e., LLM judge outputs are highly sensitive to pre-defined score ranges. We also show that similar biases exist among models from the same family. We then mitigate this bias through contrastive decoding, achieving up to 11.7% relative improvement on average in Spearman correlation with human judgments across different score ranges.

📄 PDF Abstract BibTeX arXiv:2510.18196

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models

2026-04-16 · Yanda Li, Yuhan Liu, Zirui Song, Yunchao Wei 외 arxiv

Large audio-language models (LALMs) generalize across speech, sound, and music, but unified decoders can exhibit a \emph{temporal smoothing bias}: transient acoustic cues may be underutilized in favor of temporally smoot…

Decoding Uncertainty: The Impact of Decoding Strategies for Uncertainty Estimation in Large Language Models

2025-09-20 · Wataru Hashimoto, Hidetaka Kamigaito, Taro Watanabe arxiv

Decoding strategies manipulate the probability distribution underlying the output of a language model and can therefore affect both generation quality and its uncertainty. In this study, we investigate the impact of deco…

SDCD: Structure-Disrupted Contrastive Decoding for Mitigating Hallucinations in Large Vision-Language Models

2026-01-07 · Yuxuan Xia, Siheng Wang, Peng Li arxiv

Large Vision-Language Models (LVLMs) demonstrate significant progress in multimodal understanding and reasoning, yet object hallucination remains a critical challenge. While existing research focuses on mitigating langua…

SHIELD: Suppressing Hallucinations In LVLM Encoders via Bias and Vulnerability Defense

2025-10-18 · Yiyang Huang, Liang Shi, Yitian Zhang, Yi Xu 외 arxiv

Large Vision-Language Models (LVLMs) excel in diverse cross-modal tasks. However, object hallucination, where models produce plausible but inaccurate object descriptions, remains a significant challenge. In contrast to p…

Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -

2025-04-16 · Laura Fieback, Nishilkumar Balar, Jakob Spiegelberg, Hanno Gottschalk

Despite recent advances in Large Vision Language Models (LVLMs), these models still suffer from generating hallucinatory responses that do not align with the visual input provided. To mitigate such hallucinations, we int…

Hallucination