paper-with-me

홈 › Papers

On the Trade-Off Between Transparency and Security in Adversarial Machine Learning

2025-11-14 · Lucas Fenaux, Christopher Srinivasa, Florian Kerschbaum arxiv

Transparency and security are both central to Responsible AI, but they may conflict in adversarial settings. We investigate the strategic effect of transparency for agents through the lens of transferable adversarial example attacks. In transferable adversarial example attacks, attackers maliciously perturb their inputs using surrogate models to fool a defender's target model. These models can be defended or undefended, with both players having to decide which to use. Using a large-scale empirical evaluation of nine attacks across 181 models, we find that attackers are more successful when they match the defender's decision; hence, obscurity could be beneficial to the defender. With game theory, we analyze this trade-off between transparency and security by modeling this problem as both a Nash game and a Stackelberg game, and comparing the expected outcomes. Our analysis confirms that only knowing whether a defender's model is defended or not can sometimes be enough to damage its security. This result serves as an indicator of the general trade-off between transparency and security, suggesting that transparency in AI systems can be at odds with security. Beyond adversarial machine learning, our work illustrates how game-theoretic reasoning can uncover conflicts between transparency and security.

📄 PDF Abstract BibTeX arXiv:2511.11842

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Sustainable SecureML: Quantifying Carbon Footprint of Adversarial Machine Learning

2024-03-27 · Syed Mhamudul Hasan, Abdur R. Shahid, Ahmed Imteaj

The widespread adoption of machine learning (ML) across various industries has raised sustainability concerns due to its substantial energy usage and carbon emissions. This issue becomes more pressing in adversarial ML, …

Adversarial Robustness

Learning Robust and Privacy-Preserving Representations via Information Theory

2024-12-15 · Binghui Zhang, Sayedeh Leila Noorbakhsh, Yun Dong, Yuan Hong 외

Machine learning models are vulnerable to both security attacks (e.g., adversarial examples) and privacy attacks (e.g., private attribute inference). We take the first step to mitigate both the security and privacy attac…

Adversarial RobustnessAttributePrivacy PreservingRepresentation Learning

Mitigation of Adversarial Attacks through Embedded Feature Selection

2018-08-16 · Ziyi Bao, Luis Muñoz-González, Emil C. Lupu

Machine learning has become one of the main components for task automation in many application domains. Despite the advancements and impressive achievements of machine learning, it has been shown that learning algorithms…

BIG-bench Machine Learningfeature selection

On Security and Sparsity of Linear Classifiers for Adversarial Settings

2017-08-31 · Ambra Demontis, Paolo Russu, Battista Biggio, Giorgio Fumera 외

Machine-learning techniques are widely used in security-related applications, like spam and malware detection. However, in such settings, they have been shown to be vulnerable to adversarial attacks, including the delibe…

Malware Detection

Enhancing Adversarial Robustness of IoT Intrusion Detection via SHAP-Based Attribution Fingerprinting

2025-11-09 · Dilli Prasad Sharma, Liang Xue, Xiaowei Sun, Xiaodong Lin 외 arxiv

The rapid proliferation of Internet of Things (IoT) devices has transformed numerous industries by enabling seamless connectivity and data-driven automation. However, this expansion has also exposed IoT networks to incre…

Adversarial RobustnessIntrusion Detection