paper-with-me

Papers

I-trustworthy Models. A framework for trustworthiness evaluation of probabilistic classifiers

2025-01-26 · Ritwik Vashistha, Arya Farahi

As probabilistic models continue to permeate various facets of our society and contribute to scientific advancements, it becomes a necessity to go beyond traditional metrics such as predictive accuracy and error rates and assess their trustworthiness. Grounded in the competence-based theory of trust, this work formalizes I-trustworthy framework -- a novel framework for assessing the trustworthiness of probabilistic classifiers for inference tasks by linking local calibration to trustworthiness. To assess I-trustworthiness, we use the local calibration error (LCE) and develop a method of hypothesis-testing. This method utilizes a kernel-based test statistic, Kernel Local Calibration Error (KLCE), to test local calibration of a probabilistic classifier. This study provides theoretical guarantees by offering convergence bounds for an unbiased estimator of KLCE. Additionally, we present a diagnostic tool designed to identify and measure biases in cases of miscalibration. The effectiveness of the proposed test statistic is demonstrated through its application to both simulated and real-world datasets. Finally, LCE of related recalibration methods is studied, and we provide evidence of insufficiency of existing methods to achieve I-trustworthiness.

📄 PDF Abstract BibTeX arXiv:2501.15617

Code (1)

ritwikvashistha/I-trustworthy 공식 구현

Tasks

Diagnostic

Similar Papers 제목 키워드 기반

U-Trustworthy Models.Reliability, Competence, and Confidence in Decision-Making

2024-01-04 · Ritwik Vashistha, Arya Farahi

With growing concerns regarding bias and discrimination in predictive models, the AI community has increasingly focused on assessing AI system trustworthiness. Conventionally, trustworthy AI literature relies on the prob…

Decision MakingModel SelectionPhilosophy

Automated Trustworthiness Testing for Machine Learning Classifiers

2024-06-07 · Steven Cho, Seaton Cousins-Baxter, Stefano Ruberto, Valerio Terragni

Machine Learning (ML) has become an integral part of our society, commonly used in critical domains such as finance, healthcare, and transportation. Therefore, it is crucial to evaluate not only whether ML models make co…

Word Embeddings

Probabilistic Robustness in Medical Image Classification

2026-07-04 · Yi Zhang, Siddartha Khastgir, Xingyu Zhao arxiv

Deep learning (DL) has shown strong performance in medical image classification, but its trustworthy deployment remains challenging in safety-critical clinical settings, where prediction errors under perturbations may le…

Medical Image ClassificationAdversarial Robustness

Generative Classifiers as a Basis for Trustworthy Image Classification

2020-07-29 · CVPR 2021 1 · Radek Mackowiak, Lynton Ardizzone, Ullrich Köthe, Carsten Rother

With the maturing of deep learning systems, trustworthiness is becoming increasingly important for model assessment. We understand trustworthiness as the combination of explainability and robustness. Generative classifie…

ClassificationGeneral Classificationimage-classificationImage Classification

Knowledge-Based Trust: Estimating the Trustworthiness of Web Sources

2015-02-12 · Dong Xin Luna, Gabrilovich Evgeniy, Murphy Kevin, Dang Van 외

The quality of web sources has been traditionally evaluated using exogenous signals such as the hyperlink structure of the graph. We propose a new approach that relies on endogenous signals, namely, the correctness of fa…