paper-with-me

홈 › Papers

Focus on Hiders: Exploring Hidden Threats for Enhancing Adversarial Training

2023-12-12 · CVPR 2024 1 · Qian Li, Yuxiao Hu, Yinpeng Dong, Dongxiao Zhang, Yuntian Chen

Adversarial training is often formulated as a min-max problem, however, concentrating only on the worst adversarial examples causes alternating repetitive confusion of the model, i.e., previously defended or correctly classified samples are not defensible or accurately classifiable in subsequent adversarial training. We characterize such non-ignorable samples as "hiders", which reveal the hidden high-risk regions within the secure area obtained through adversarial training and prevent the model from finding the real worst cases. We demand the model to prevent hiders when defending against adversarial examples for improving accuracy and robustness simultaneously. By rethinking and redefining the min-max optimization problem for adversarial training, we propose a generalized adversarial training algorithm called Hider-Focused Adversarial Training (HFAT). HFAT introduces the iterative evolution optimization strategy to simplify the optimization problem and employs an auxiliary model to reveal hiders, effectively combining the optimization directions of standard adversarial training and prevention hiders. Furthermore, we introduce an adaptive weighting mechanism that facilitates the model in adaptively adjusting its focus between adversarial examples and hiders during different training periods. We demonstrate the effectiveness of our method based on extensive experiments, and ensure that HFAT can provide higher robustness and accuracy.

📄 PDF Abstract BibTeX arXiv:2312.07067

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Possible traces of resonance signaling in the genome

2019-10-25 · Ivan Savelev, Max Myakishev-Rempel

Although theories regarding the role of sequence-specific DNA resonance in biology have abounded for over 40 years, the published evidence for it is lacking. Here, the authors reasoned that for sustained resonance signal…

Replication of Multi-agent Reinforcement Learning for the "Hide and Seek" Problem

2023-10-09 · Haider Kamal, Muaz A. Niazi, Hammad Afzal

Reinforcement learning generates policies based on reward functions and hyperparameters. Slight changes in these can significantly affect results. The lack of documentation and reproducibility in Reinforcement learning r…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning

Exploring Multi-Level Threats in Telegram Data with AI-Human Annotation: A Preliminary Study

2023-12-15 · 2023 22nd IEEE International Conference on Machine Learning and Applications (ICMLA) 2023 12 · Kamalakkannan Ravi, Adan Ernesto Vela, Elizabeth Jenaway, Steven Windisch

This research addresses the crucial challenge of effectively measuring threats in social media comments targeting voting, public officials, and institutions in the United States. Our understanding of these online threats…

Information RetrievalText ClassificationTransfer LearningViolence and Weaponized Violence Detection

Uncovering the Hidden Threat of Text Watermarking from Users with Cross-Lingual Knowledge

2025-02-23 · Mansour Al Ghanim, Jiaqi Xue, Rochana Prih Hastuti, Mengxin Zheng 외

In this study, we delve into the hidden threats posed to text watermarking by users with cross-lingual knowledge. While most research focuses on watermarking methods for English, there is a significant gap in evaluating …

Exploring Prompt Engineering: A Systematic Review with SWOT Analysis

2024-10-09 · Aditi Singh, Abul Ehtesham, Gaurav Kumar Gupta, Nikhil Kumar Chatta 외

In this paper, we conduct a comprehensive SWOT analysis of prompt engineering techniques within the realm of Large Language Models (LLMs). Emphasizing linguistic principles, we examine various techniques to identify thei…

Language ModelingLanguage ModellingPrompt Engineering