paper-with-me

홈 › Papers

Improving the Reliability of Large Language Models by Leveraging Uncertainty-Aware In-Context Learning

2023-10-07 · Yuchen Yang, Houqiang Li, Yanfeng Wang, Yu Wang

In recent years, large-scale language models (LLMs) have gained attention for their impressive text generation capabilities. However, these models often face the challenge of "hallucination," which undermines their reliability. In this study, we introduce an uncertainty-aware in-context learning framework to empower the model to enhance or reject its output in response to uncertainty. Human-defined methods for estimating uncertainty typically assume that "uncertainty is lower when the model's response is correct compared to when it is incorrect." However, setting a precise threshold to distinguish correctness is challenging. Therefore, we introduce uncertainty information as an intermediary variable that implicitly influences the model's behavior. Our innovative uncertainty-aware in-context learning framework involves fine-tuning the LLM using a calibration dataset. Our aim is to improve the model's responses by filtering out answers with high uncertainty while considering the model's knowledge limitations. We evaluate the model's knowledge by examining multiple responses to the same question for the presence of a correct answer. When the model lacks relevant knowledge, the response should indicate that the question cannot be answered. Conversely, when the model has relevant knowledge, the response should provide the correct answer. Extensive experiments confirm the effectiveness of our framework, leading to two key findings. First, the logit output values of the LLM partly reflect inherent uncertainty. Second, our model autonomously recognizes uncertainty, resulting in improved responses.

📄 PDF Abstract BibTeX arXiv:2310.04782

Code (0)

등록된 구현이 없습니다.

Tasks

HallucinationIn-Context LearningText Generation

Similar Papers 제목 키워드 기반

Uncertainty-Aware Step-wise Verification with Generative Reward Models

2025-02-16 · Zihuiwen Ye, Luckeciano Carvalho Melo, Younesse Kaddar, Phil Blunsom 외

Complex multi-step reasoning tasks, such as solving mathematical problems, remain challenging for large language models (LLMs). While outcome supervision is commonly used, process supervision via process reward models (P…

Mathematical ReasoningUncertainty Quantification

Uncertainty-Aware Adaptation of Large Language Models for Protein-Protein Interaction Analysis

2025-02-10 · Sanket Jantre, Tianle Wang, Gilchan Park, Kriti Chopra 외

Identification of protein-protein interactions (PPIs) helps derive cellular mechanistic understanding, particularly in the context of complex conditions such as neurodegenerative disorders, metabolic syndromes, and cance…

Uncertainty Quantification

Uncertainty Quantification of Large Language Models through Multi-Dimensional Responses

2025-02-24 · Tiejin Chen, Longchao Da, Xiaoou Liu, Vagelis Papalexakis 외

Large Language Models (LLMs) have demonstrated remarkable capabilities across various tasks due to large training datasets and powerful transformer architecture. However, the reliability of responses from LLMs remains a …

Decision MakingSemantic SimilaritySemantic Textual SimilarityTensor Decomposition+1

Can LLMs Detect Their Confabulations? Estimating Reliability in Uncertainty-Aware Language Models

2025-08-11 · Tianyi Zhou, Johanne Medina, Sanjay Chawla arxiv

Large Language Models (LLMs) are prone to generating fluent but incorrect content, known as confabulation, which poses increasing risks in multi-turn or agentic applications where outputs may be reused as context. In thi…

Probabilistic Machine Learning for Noisy Labels in Earth Observation

2025-04-04 · Spyros Kondylatos, Nikolaos Ioannis Bountos, Ioannis Prapas, Angelos Zavras 외

Label noise poses a significant challenge in Earth Observation (EO), often degrading the performance and reliability of supervised Machine Learning (ML) models. Yet, given the critical nature of several EO applications, …

Earth ObservationUncertainty Quantification