paper-with-me

Papers

Towards Certification of Uncertainty Calibration under Adversarial Attacks

2024-05-22 · Cornelius Emde, Francesco Pinto, Thomas Lukasiewicz, Philip H. S. Torr, Adel Bibi

Since neural classifiers are known to be sensitive to adversarial perturbations that alter their accuracy, \textit{certification methods} have been developed to provide provable guarantees on the insensitivity of their predictions to such perturbations. Furthermore, in safety-critical applications, the frequentist interpretation of the confidence of a classifier (also known as model calibration) can be of utmost importance. This property can be measured via the Brier score or the expected calibration error. We show that attacks can significantly harm calibration, and thus propose certified calibration as worst-case bounds on calibration under adversarial perturbations. Specifically, we produce analytic bounds for the Brier score and approximate bounds via the solution of a mixed-integer program on the expected calibration error. Finally, we propose novel calibration attacks and demonstrate how they can improve model calibration through \textit{adversarial calibration training}.

📄 PDF Abstract BibTeX arXiv:2405.13922

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Et Tu Certifications: Robustness Certificates Yield Better Adversarial Examples

2023-02-09 · Andrew C. Cullen, Shijie Liu, Paul Montague, Sarah M. Erfani 외

In guaranteeing the absence of adversarial examples in an instance's neighbourhood, certification mechanisms play an important role in demonstrating neural net robustness. In this paper, we ask if these certifications ca…

PROSAC: Provably Safe Certification for Machine Learning Models under Adversarial Attacks

2024-02-04 · Chen Feng, Ziquan Liu, Zhuo Zhi, Ilija Bogunovic 외

It is widely known that state-of-the-art machine learning models, including vision and language models, can be seriously compromised by adversarial perturbations. It is therefore increasingly relevant to develop capabili…

Adversarial AttackBayesian Optimization

Uncertainty Quantification for Collaborative Object Detection Under Adversarial Attacks

2025-02-04 · Huiqun Huang, Cong Chen, Jean-Philippe Monteuuis, Jonathan Petit 외

Collaborative Object Detection (COD) and collaborative perception can integrate data or features from various entities, and improve object detection accuracy compared with individual perception. However, adversarial atta…

Adversarial RobustnessAutonomous DrivingAutonomous VehiclesConformal Prediction+4

Fast Adversarial Robustness Certification of Nearest Prototype Classifiers for Arbitrary Seminorms

2020-12-01 · NeurIPS 2020 12 · Sascha Saralajew, Lars Holdijk, Thomas Villmann

Methods for adversarial robustness certification aim to provide an upper bound on the test error of a classifier under adversarial manipulation of its input. Current certification methods are computationally expensive an…

Adversarial RobustnessQuantizationTriplet

Calibrating Uncertainty for Zero-Shot Adversarial CLIP

2025-12-15 · Wenjing Lu, Zerui Tao, Yuning Qiu, Dongping Zhang 외 arxiv

CLIP delivers strong zero-shot classification but remains highly vulnerable to adversarial attacks. Prior adversarial fine-tuning work primarily matches predicted logits between clean and adversarial examples, which over…

Zero-shot GeneralizationAdversarial Robustness