paper-with-me

홈 › Papers

Confidence Matters: Uncertainty Quantification and Precision Assessment of Deep Learning-based CMR Biomarker Estimates Using Scan-rescan Data

2026-03-25 · Dewmini Hasara Wickremasinghe, Michelle Gibogwe, Andrew Bell, Esther Puyol-Antón, Muhummad Sohaib Nazir, Reza Razavi, Bruno Paun, Paul Aljabar, Andrew P. King arxiv

The performance of deep learning (DL) methods for the analysis of cine cardiovascular magnetic resonance (CMR) is typically assessed in terms of accuracy, overlooking precision. In this work, uncertainty estimation techniques, namely deep ensemble, test-time augmentation, and Monte Carlo dropout, are applied to a state-of-the-art DL pipeline for cardiac functional biomarker estimation, and new distribution-based metrics are proposed for the assessment of biomarker precision. The model achieved high accuracy (average Dice 87%) and point estimate precision on two external validation scan-rescan CMR datasets. However, distribution-based metrics showed that the overlap between scan/rescan confidence intervals was >50% in less than 45% of the cases. Statistical similarity tests between scan and rescan biomarkers also resulted in significant differences for over 65% of the cases. We conclude that, while point estimate metrics might suggest good performance, distributional analyses reveal lower precision, highlighting the need to use more representative metrics to assess scan-rescan agreement.

📄 PDF Abstract BibTeX arXiv:2603.26789

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Estimating prevalence with precision and accuracy

2025-07-08 · Aime Bienfait Igiraneza, Christophe Fraser, Robert Hinch

Unlike classification, whose goal is to estimate the class of each data point in a dataset, prevalence estimation or quantification is a task that aims to estimate the distribution of classes in a dataset. The two main t…

Uncertainty Quantification

Label-Confidence-Aware Uncertainty Estimation in Natural Language Generation

2024-12-10 · Qinhong Lin, Linna Zhou, Zhongliang Yang, Yuang Cai

Large Language Models (LLMs) display formidable capabilities in generative tasks but also pose potential risks due to their tendency to generate hallucinatory responses. Uncertainty Quantification (UQ), the evaluation of…

Text GenerationUncertainty Quantification

Efficient Self-Evaluation for Diffusion Language Models via Sequence Regeneration

2026-03-03 · Linhao Zhong, Linyu Wu, Wen Wang, Yuling Xi 외 arxiv

Diffusion large language models (dLLMs) have recently attracted significant attention for their ability to enhance diversity, controllability, and parallelism. However, their non-sequential, bidirectionally masked genera…

Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models

2026-06-02 · Qi Cao, Takeshi Kojima, Andrew Gambardella, Helinyi Peng 외 arxiv

Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while being factually incorrect. This problem is compounded by the lack of exp…

Generating with Confidence: Uncertainty Quantification for Black-box Large Language Models

2023-05-30 · Zhen Lin, Shubhendu Trivedi, Jimeng Sun

Large language models (LLMs) specializing in natural language generation (NLG) have recently started exhibiting promising capabilities across a variety of domains. However, gauging the trustworthiness of responses genera…

ManagementQuestion AnsweringText GenerationUncertainty Quantification