paper-with-me

홈 › Papers

A unifying Bayesian framework for adversarial robustness

2025-10-10 · Pablo G. Arce, Roi Naveiro, David Ríos Insua arxiv

The vulnerability of machine learning models to adversarial attacks remains a critical societal security challenge. Traditional defenses, such as adversarial training, typically robustify models by minimizing a worst-case loss. These deterministic approaches do not account for uncertainty in the adversary's attack. While stochastic defenses placing a probability distribution on the adversary exist, they often lack statistical rigor and fail to make explicit their underlying assumptions. To resolve these issues, we introduce a formal Bayesian framework that models adversarial uncertainty through a stochastic channel, articulating all probabilistic assumptions. This yields two robustification strategies: a proactive defense enacted during training, aligned with adversarial training, and a reactive defense enacted during operations, aligned with adversarial purification. Several state-of-the-art defenses can be recovered as limiting cases of our model. We empirically validate our methodology, showcasing the benefits of explicitly modeling adversarial uncertainty.

📄 PDF Abstract BibTeX arXiv:2510.09288

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial Robustness

Similar Papers 제목 키워드 기반

Unifying Adversarial Robustness and Training Across Text Scoring Models

2026-01-31 · Manveer Singh Tamber, Hosna Oyarhoseini, Jimmy Lin arxiv

Research on adversarial robustness in language models is currently fragmented across applications and attacks, obscuring shared vulnerabilities. In this work, we propose unifying the study of adversarial robustness in te…

Adversarial Robustness

Robust Bayesian Neural Networks by Spectral Expectation Bound Regularization

2021-06-19 · CVPR 2021 1 · Jiaru Zhang, Yang Hua, Zhengui Xue, Tao Song 외

Bayesian neural networks have been widely used in many applications because of the distinctive probabilistic representation framework. Even though Bayesian neural networks have been found more robust to adversarial a…

Generalization Certificates for Adversarially Robust Bayesian Linear Regression

2025-02-20 · Mahalakshmi Sabanayagam, Russell Tsuchida, Cheng Soon Ong, Debarghya Ghoshdastidar

Adversarial robustness of machine learning models is critical to ensuring reliable performance under data perturbations. Recent progress has been on point estimators, and this paper considers distributional predictors. F…

Adversarial RobustnessBayesian Inferenceregression

A PAC-Bayes Analysis of Adversarial Robustness

2021-02-19 · NeurIPS 2021 12 · Paul Viallard, Guillaume Vidot, Amaury Habrard, Emilie Morvant

We propose the first general PAC-Bayesian generalization bounds for adversarial robustness, that estimate, at test time, how much a model will be invariant to imperceptible perturbations in the input. Instead of deriving…

Adversarial RobustnessGeneralization Boundsvalid

Contributions to Large Scale Bayesian Inference and Adversarial Machine Learning

2021-09-25 · Víctor Gallego

The rampant adoption of ML methodologies has revealed that models are usually adopted to make decisions without taking into account the uncertainties in their predictions. More critically, they can be vulnerable to adver…

Bayesian InferenceBIG-bench Machine LearningTime Series Analysis