paper-with-me

홈 › Papers

Unsupervised Conformal Inference: Bootstrapping and Alignment to Control LLM Uncertainty

2025-09-26 · Lingyou Pang, Lei Huang, Jianyu Lin, Tianyu Wang, Akira Horiguchi, Alexander Aue, Carey E. Priebe arxiv

Deploying black-box LLMs requires managing uncertainty in the absence of token-level probability or true labels. We propose introducing an unsupervised conformal inference framework for generation, which integrates: generative models, incorporating: (i) an LLM-compatible atypical score derived from response-embedding Gram matrix, (ii) UCP combined with a bootstrapping variant (BB-UCP) that aggregates residuals to refine quantile precision while maintaining distribution-free, finite-sample coverage, and (iii) conformal alignment, which calibrates a single strictness parameter $τ$ so a user predicate (e.g., factuality lift) holds on unseen batches with probability $\ge 1-α$. Across different benchmark datasets, our gates achieve close-to-nominal coverage and provide tighter, more stable thresholds than split UCP, while consistently reducing the severity of hallucination, outperforming lightweight per-response detectors with similar computational demands. The result is a label-free, API-compatible gate for test-time filtering that turns geometric signals into calibrated, goal-aligned decisions.

📄 PDF Abstract BibTeX arXiv:2509.23002

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Constructing Prediction Intervals with Neural Networks: An Empirical Evaluation of Bootstrapping and Conformal Inference Methods

2022-10-07 · Alex Contarino, Christine Schubert Kabban, Chancellor Johnstone, Fairul Mohd-Zaid

Artificial neural networks (ANNs) are popular tools for accomplishing many machine learning tasks, including predicting continuous outcomes. However, the general lack of confidence measures provided with ANN predictions …

Prediction Intervals

Reliable Inference in Edge-Cloud Model Cascades via Conformal Alignment

2025-10-20 · Jiayi Huang, Sangwoo Park, Nicola Paoletti, Osvaldo Simeone arxiv

Edge intelligence enables low-latency inference via compact on-device models, but assuring reliability remains challenging. We study edge-cloud cascades that must preserve conditional coverage: whenever the edge returns …

Image Classification

Optimized Conformal Selection: Powerful Selective Inference After Conformity Score Optimization

2024-11-27 · Tian Bai, Ying Jin

Model selection/optimization in conformal inference is challenging, since it may break the exchangeability between labeled and unlabeled data. We study this problem in the context of conformal selection, which uses confo…

Drug DiscoveryModel OptimizationModel Selectionvalid

Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment

2026-01-24 · Tiejin Chen, Xiaoou Liu, Vishnu Nandam, Kuan-Ru Liou 외 arxiv

Preference-based alignment like Reinforcement Learning from Human Feedback (RLHF) learns from pairwise preferences, yet the labels are often noisy and inconsistent. Existing uncertainty-aware approaches weight preference…

Reinforcement Learning

Denoised Conformal Alignment for Reliable Selection of Conditional Average Treatment Effect Predictions

2026-07-03 · Xinyun Lu, Haoang Chi, Zhiheng Zhang arxiv

In selective deployment, practitioners act only on a model-chosen subset of individuals based on predicted conditional average treatment effects, but marginal conformal guarantees need not control reliability on that sel…