paper-with-me

Papers

Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge

2025-05-18 · Luyu Chen, Zeyu Zhang, Haoran Tan, Quanyu Dai, Hao Yang, Zhenhua Dong, Xu Chen

LLMs have emerged as powerful evaluators in the LLM-as-a-Judge paradigm, offering significant efficiency and flexibility compared to human judgments. However, previous methods primarily rely on single-point evaluations, overlooking the inherent diversity and uncertainty in human evaluations. This approach leads to information loss and decreases the reliability of evaluations. To address this limitation, we propose a novel training framework that explicitly aligns the LLM-generated judgment distribution with empirical human distributions. Specifically, we propose a distributional alignment objective based on KL divergence, combined with an auxiliary cross-entropy regularization to stabilize the training process. Furthermore, considering that empirical distributions may derive from limited human annotations, we incorporate adversarial training to enhance model robustness against distribution perturbations. Extensive experiments across various LLM backbones and evaluation tasks demonstrate that our framework significantly outperforms existing closed-source LLMs and conventional single-point alignment methods, with improved alignment quality, evaluation accuracy, and robustness.

📄 PDF Abstract BibTeX arXiv:2505.12301

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

Beyond Consensus: Downward Bias and Role Asymmetry in Multi-Agent LLM Judges for Subjective Evaluation

2026-08-31 · Minsoo Song, Chanwoo Kim, Sugyeong Eo, Chanjun Park arxiv

Multi-Agent Debate (MAD) has been widely adopted to improve LLM-based evaluation by prompting multiple agents to negotiate and reach a consensus. However, for subjective rubric-based scoring, inter-agent agreement does n…

Principles Do Not Apply Themselves: A Hermeneutic Perspective on AI Alignment

2026-04-12 · Behrooz Razeghi arxiv

AI alignment is often framed as the task of ensuring that an AI system follows a set of stated principles or human preferences, but general principles rarely determine their own application in concrete cases. When princi…

Beyond the Surface: Enhancing LLM-as-a-Judge Alignment with Human via Internal Representations

2025-08-05 · Peng Lai, Jianjie Zheng, Sijie Cheng, Yun Chen 외 arxiv

The growing scale of evaluation tasks has led to the widespread adoption of automated evaluation using LLMs, a paradigm known as "LLM-as-a-judge". However, improving its alignment with human preferences without complex p…

The Pluralistic Moral Gap: Understanding Judgment and Value Differences between Humans and Large Language Models

2025-07-23 · Giuseppe Russo, Debora Nozza, Paul Röttger, Dirk Hovy arxiv

People increasingly rely on Large Language Models (LLMs) for moral advice, which may influence humans' decisions. Yet, little is known about how closely LLMs align with human moral judgments. To address this, we introduc…

MSD-Score: Multi-Scale Distributional Scoring for Reference-Free Image Caption Evaluation

2026-05-07 · Shichao Kan, Xuyang Zhang, Haojie Zhang, Zhe Zhu 외 arxiv

Evaluating image captions without references remains challenging because global embedding similarity often misses fine-grained mismatches such as hallucinated objects, missing attributes, or incorrect relations. We propo…

Image-text matching