paper-with-me

홈 › Papers

Defense Through Diverse Directions

2020-03-24 · ICML 2020 1 · Christopher M. Bender, Yang Li, Yifeng Shi, Michael K. Reiter, Junier B. Oliva

In this work we develop a novel Bayesian neural network methodology to achieve strong adversarial robustness without the need for online adversarial training. Unlike previous efforts in this direction, we do not rely solely on the stochasticity of network weights by minimizing the divergence between the learned parameter distribution and a prior. Instead, we additionally require that the model maintain some expected uncertainty with respect to all input covariates. We demonstrate that by encouraging the network to distribute evenly across inputs, the network becomes less susceptible to localized, brittle features which imparts a natural robustness to targeted perturbations. We show empirical robustness on several benchmark datasets.

📄 PDF Abstract BibTeX arXiv:2003.10602

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial Robustness

Similar Papers 제목 키워드 기반

Toward a Generalized Defense Across Sparse, Continuous, and Structured Parameter Attacks

2026-06-03 · Bin Duan, Zeyu Bai, Guowei Yang arxiv

Deep neural networks are increasingly deployed across heterogeneous and partially untrusted environments, where models are distributed through cloud storage, CI/CD pipelines, containerized services, and edge execution pl…

Privacy in Fine-tuning Large Language Models: Attacks, Defenses, and Future Directions

2024-12-21 · Hao Du, Shang Liu, Lele Zheng, Yang Cao 외

Fine-tuning has emerged as a critical process in leveraging Large Language Models (LLMs) for specific downstream tasks, enabling these models to achieve state-of-the-art performance across various domains. However, the f…

Federated LearningPrivacy Preserving

Defense Against the Dark Arts: An overview of adversarial example security research and future research directions

2018-06-11 · Ian Goodfellow

This article presents a summary of a keynote lecture at the Deep Learning Security workshop at IEEE Security and Privacy 2018. This lecture summarizes the state of the art in defenses against adversarial examples and pro…

Deep Learning

Distilling the Undistillable: Learning from a Nasty Teacher

2022-10-21 · Surgan Jandial, Yash Khasbage, Arghya Pal, Vineeth N Balasubramanian 외

The inadvertent stealing of private/sensitive information using Knowledge Distillation (KD) has been getting significant attention recently and has guided subsequent defense efforts considering its critical nature. Recen…

Knowledge Distillation

Backdoor Attacks and Defenses in Federated Learning: Survey, Challenges and Future Research Directions

2023-03-03 · Thuy Dung Nguyen, Tuan Nguyen, Phi Le Nguyen, Hieu H. Pham 외

Federated learning (FL) is a machine learning (ML) approach that allows the use of distributed data without compromising personal privacy. However, the heterogeneous distribution of data among clients in FL can make it d…

Backdoor AttackFederated LearningSurvey