paper-with-me

Papers

Adaptation with Self-Evaluation to Improve Selective Prediction in LLMs

2023-10-18 · Jiefeng Chen, Jinsung Yoon, Sayna Ebrahimi, Sercan O Arik, Tomas Pfister, Somesh Jha

Large language models (LLMs) have recently shown great advances in a variety of tasks, including natural language understanding and generation. However, their use in high-stakes decision-making scenarios is still limited due to the potential for errors. Selective prediction is a technique that can be used to improve the reliability of the LLMs by allowing them to abstain from making predictions when they are unsure of the answer. In this work, we propose a novel framework for adaptation with self-evaluation to improve the selective prediction performance of LLMs. Our framework is based on the idea of using parameter-efficient tuning to adapt the LLM to the specific task at hand while improving its ability to perform self-evaluation. We evaluate our method on a variety of question-answering (QA) datasets and show that it outperforms state-of-the-art selective prediction methods. For example, on the CoQA benchmark, our method improves the AUACC from 91.23% to 92.63% and improves the AUROC from 74.61% to 80.25%.

📄 PDF Abstract BibTeX arXiv:2310.11689

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingNatural Language UnderstandingPredictionQuestion Answering

Similar Papers 제목 키워드 기반

Conformal Uncertainty Indicator for Continual Test-Time Adaptation

2025-02-05 · Fan Lyu, Hanyu Zhao, Ziqi Shi, Ye Liu 외

Continual Test-Time Adaptation (CTTA) aims to adapt models to sequentially changing domains during testing, relying on pseudo-labels for self-adaptation. However, incorrect pseudo-labels can accumulate, leading to perfor…

Conformal PredictionTest-time Adaptation

AUGCO: Augmentation Consistency-guided Self-training for Source-free Domain Adaptive Semantic Segmentation

2021-07-21 · Viraj Prabhu, Shivam Khare, Deeksha Kartik, Judy Hoffman

Most modern approaches for domain adaptive semantic segmentation rely on continued access to source data during adaptation, which may be infeasible due to computational or privacy constraints. We focus on source-free dom…

Domain AdaptationSegmentationSemantic SegmentationSource-Free Domain Adaptation

Self-Evaluation Improves Selective Generation in Large Language Models

2023-12-14 · Jie Ren, Yao Zhao, Tu Vu, Peter J. Liu 외

Safe deployment of large language models (LLMs) may benefit from a reliable method for assessing their generated content to determine when to abstain or to selectively generate. While likelihood-based metrics such as per…

Multiple-choiceTruthfulQA

Uncertainty-Aware Last-Layer Adaptation of RETFound for Referable Diabetic Retinopathy Screening Under Dataset Shift

2026-06-30 · Karim Mardhani arxiv

This paper presents a safety-centered empirical evaluation of uncertainty-aware last-layer adaptation for referable diabetic retinopathy screening using RETFound, a self-supervised vision-transformer retinal foundation m…

ORB-SfMLearner: ORB-Guided Self-supervised Visual Odometry with Selective Online Adaptation

2024-09-18 · Yanlin Jin, Rui-Yang Ju, Haojun Liu, Yuzhong Zhong

Deep visual odometry, despite extensive research, still faces limitations in accuracy and generalizability that prevent its broader application. To address these challenges, we propose an Oriented FAST and Rotated BRIEF …

Motion EstimationVisual Odometry