paper-with-me

Papers

UNCERTAINTY-LINE: Length-Invariant Estimation of Uncertainty for Large Language Models

2025-05-25 · Roman Vashurin, Maiya Goloburda, Preslav Nakov, Maxim Panov

Large Language Models (LLMs) have become indispensable tools across various applications, making it more important than ever to ensure the quality and the trustworthiness of their outputs. This has led to growing interest in uncertainty quantification (UQ) methods for assessing the reliability of LLM outputs. Many existing UQ techniques rely on token probabilities, which inadvertently introduces a bias with respect to the length of the output. While some methods attempt to account for this, we demonstrate that such biases persist even in length-normalized approaches. To address the problem, here we propose UNCERTAINTY-LINE: (Length-INvariant Estimation), a simple debiasing procedure that regresses uncertainty scores on output length and uses the residuals as corrected, length-invariant estimates. Our method is post-hoc, model-agnostic, and applicable to a range of UQ measures. Through extensive evaluation on machine translation, summarization, and question-answering tasks, we demonstrate that UNCERTAINTY-LINE: consistently improves over even nominally length-normalized UQ methods uncertainty estimates across multiple metrics and models.

📄 PDF Abstract BibTeX arXiv:2505.19060

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationQuestion AnsweringUncertainty Quantification

Similar Papers 제목 키워드 기반

Improving the Performance of Robust Control through Event-Triggered Learning

2022-07-28 · Alexander von Rohr, Friedrich Solowjow, Sebastian Trimpe

Robust controllers ensure stability in feedback loops designed under uncertainty but at the cost of performance. Model uncertainty in time-invariant systems can be reduced by recently proposed learning-based methods, whi…

Incorporating Uncertainty from Speaker Embedding Estimation to Speaker Verification

2023-02-23 · Qiongqiong Wang, Kong Aik Lee, Tianchi Liu

Speech utterances recorded under differing conditions exhibit varying degrees of confidence in their embedding estimates, i.e., uncertainty, even if they are extracted using the same neural network. This paper aims to in…

Speaker Verification

Semantic Reformulation Entropy for Robust Hallucination Detection in QA Tasks

2025-09-22 · Chaodong Tong, Qi Zhang, Lei Jiang, Yanbing Liu 외 arxiv

Reliable question answering with large language models (LLMs) is challenged by hallucinations, fluent but factually incorrect outputs arising from epistemic uncertainty. Existing entropy-based semantic-level uncertainty …

Question Answering

Sample Complexity of the Robust LQG Regulator with Coprime Factors Uncertainty

2021-09-29 · Yifei Zhang, Sourav Kumar Ukil, Ephraim Neimand, Serban Sabau 외

This paper addresses the end-to-end sample complexity bound for learning the H2 optimal controller (the Linear Quadratic Gaussian (LQG) problem) with unknown dynamics, for potentially unstable Linear Time Invariant (LTI)…

Time SeriesTime Series Analysis

Learning-Enhanced Observer for Linear Time-Invariant Systems with Parametric Uncertainty

2025-11-20 · Hao Shu arxiv

This work introduces a learning-enhanced observer (LEO) for linear time-invariant systems with uncertain dynamics. Rather than relying solely on nominal models, the proposed framework treats the system matrices as optimi…