paper-with-me

홈 › Papers

LLMs as Implicit Imputers: Uncertainty Should Scale with Missing Information

2026-05-13 · Stef van Buuren arxiv

Large language models (LLMs) are increasingly deployed in settings where the available context is incomplete or degraded. We argue that an LLM generating answers under incomplete context can be viewed as an implicit imputer, and evaluated against a criterion from the multiple imputation (MI) literature: uncertainty should scale with the amount of missing information. We assess this criterion on SQuAD, using a controlled framework in which context availability is varied across five levels. We evaluate two answer-level uncertainty measures that can be estimated from repeated sampling: sampling-based confidence (empirical mode frequency) and response entropy. Confidence fails to reflect increasing missingness: it remains high even as accuracy collapses. Entropy, by contrast, increases with context removal, consistent with the MI analogy, and explains substantially more variance in accuracy than confidence across all evidence levels (quadratic $R^2$ gap up to 0.057). We further introduce a black-box diagnostic $ρ_R(α)$ that estimates the proportion of baseline uncertainty resolved by context level $α$, requiring only repeated sampling with and without context. These results suggest that entropy is a more responsive black-box uncertainty measure than confidence under incomplete context.

📄 PDF Abstract BibTeX arXiv:2605.13188

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Missingness-aware Data Imputation via AI-powered Bayesian Generative Modeling

2026-05-03 · Qiao Liu arxiv

Missing data imputation remains a fundamental challenge in modern data science, especially when uncertainty quantification is essential. In this work, we propose MissBGM, an AI-powered missing data imputation method via …

Stochastic OptimizationBayesian Inference

Improving the Reliability of Large Language Models by Leveraging Uncertainty-Aware In-Context Learning

2023-10-07 · Yuchen Yang, Houqiang Li, Yanfeng Wang, Yu Wang

In recent years, large-scale language models (LLMs) have gained attention for their impressive text generation capabilities. However, these models often face the challenge of "hallucination," which undermines their relia…

HallucinationIn-Context LearningText Generation

Beyond Accuracy: An Empirical Study of Uncertainty Estimation in Imputation

2025-11-26 · Zarin Tahia Hossain, Mostafa Milani arxiv

Handling missing data is a central challenge in data-driven analysis. Modern imputation methods not only aim for accurate reconstruction but also differ in how they represent and quantify uncertainty. Yet, the reliabilit…

Curriculum-Aware Interpolate-then-Refine: Learned Physiological Time-Series Imputation under Realistic Missingness

2026-08-21 · Yu-Chao Huang, Haochen Zhang, Nicholas Konz, Tianlong Chen arxiv

Imputing physiological time series (arterial blood pressure, blood glucose, etc.) is essential for addressing the missingness that pervades clinical data. Yet modern imputation methods perform poorly in this domain: a re…

D-IF: Uncertainty-aware Human Digitization via Implicit Distribution Field

2023-08-17 · ICCV 2023 1 · Xueting Yang, Yihao Luo, Yuliang Xiu, Wei Wang 외

Realistic virtual humans play a crucial role in numerous industries, such as metaverse, intelligent healthcare, and self-driving simulation. But creating them on a large scale with high levels of realism remains a challe…