paper-with-me

홈 › Papers

Evaluating the Robustness of Bayesian Neural Networks Against Different Types of Attacks

2021-06-17 · Yutian Pang, Sheng Cheng, Jueming Hu, Yongming Liu

To evaluate the robustness gain of Bayesian neural networks on image classification tasks, we perform input perturbations, and adversarial attacks to the state-of-the-art Bayesian neural networks, with a benchmark CNN model as reference. The attacks are selected to simulate signal interference and cyberattacks towards CNN-based machine learning systems. The result shows that a Bayesian neural network achieves significantly higher robustness against adversarial attacks generated against a deterministic neural network model, without adversarial training. The Bayesian posterior can act as the safety precursor of ongoing malicious activities. Furthermore, we show that the stochastic classifier after the deterministic CNN extractor has sufficient robustness enhancement rather than a stochastic feature extractor before the stochastic classifier. This advises on utilizing stochastic layers in building decision-making pipelines within a safety-critical domain.

📄 PDF Abstract BibTeX arXiv:2106.09223

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Makingimage-classificationImage Classification

Similar Papers 제목 키워드 기반

Transfer of Adversarial Robustness Between Perturbation Types

2019-05-03 · Daniel Kang, Yi Sun, Tom Brown, Dan Hendrycks 외

We study the transfer of adversarial robustness of deep neural networks between different perturbation types. While most work on adversarial examples has focused on $L_\infty$ and $L_2$-bounded perturbations, these do no…

Adversarial Robustness

Make an Offer They Can't Refuse: Grounding Bayesian Persuasion in Real-World Dialogues without Pre-Commitment

2025-10-15 · Buwei He, Yang Liu, Zhaowei Zhang, Zixia Jia 외 arxiv

Large language models (LLMs) still struggle with strategic persuasion, largely because existing approaches either neglect information asymmetry or rely on unrealistic pre-commitment assumptions. We introduce a type-induc…

Perturbation Type Categorization for Multiple $\ell_p$ Bounded Adversarial Robustness

2021-01-01 · Pratyush Maini, Xinyun Chen, Bo Li, Dawn Song

Despite the recent advances in $\textit{adversarial training}$ based defenses, deep neural networks are still vulnerable to adversarial attacks outside the perturbation type they are trained to be robust against. Recent …

Adversarial RobustnessVocal Bursts Type Prediction

On Evaluating the Adversarial Robustness of Semantic Segmentation Models

2023-06-25 · Levente Halmosi, Mark Jelasity

Achieving robustness against adversarial input perturbation is an important and intriguing problem in machine learning. In the area of semantic image segmentation, a number of adversarial training approaches have been pr…

Adversarial Robustnessimage-classificationImage ClassificationImage Segmentation+1

Truth or Sophistry? LoFa: A Benchmark for LLM Robustness Against Logical Fallacies

2026-06-30 · Xudong Shen, Li Yuan, Ye Chen, Xin Wu 외 arxiv

Large Language Models (LLMs) exhibit strong semantic capabilities, yet their resilience to manipulative linguistic patterns such as logical fallacies remains underexplored. Prior work has primarily examined whether LLMs …

Logical Fallacies