paper-with-me

Papers

Offensive Security for AI Systems: Concepts, Practices, and Applications

2025-05-09 · Josh Harguess, Chris M. Ward

As artificial intelligence (AI) systems become increasingly adopted across sectors, the need for robust, proactive security strategies is paramount. Traditional defensive measures often fall short against the unique and evolving threats facing AI-driven technologies, making offensive security an essential approach for identifying and mitigating risks. This paper presents a comprehensive framework for offensive security in AI systems, emphasizing proactive threat simulation and adversarial testing to uncover vulnerabilities throughout the AI lifecycle. We examine key offensive security techniques, including weakness and vulnerability assessment, penetration testing, and red teaming, tailored specifically to address AI's unique susceptibilities. By simulating real-world attack scenarios, these methodologies reveal critical insights, informing stronger defensive strategies and advancing resilience against emerging threats. This framework advances offensive AI security from theoretical concepts to practical, actionable methodologies that organizations can implement to strengthen their AI systems against emerging threats.

📄 PDF Abstract BibTeX arXiv:2505.06380

Code (0)

등록된 구현이 없습니다.

Tasks

Red Teaming

Similar Papers 제목 키워드 기반

A Survey on Offensive AI Within Cybersecurity

2024-09-26 · Sahil Girhepuje, Aviral Verma, Gaurav Raina

Artificial Intelligence (AI) has witnessed major growth and integration across various domains. As AI systems become increasingly prevalent, they also become targets for threat actors to manipulate their functionality fo…

Survey

Benchmarking Practices in LLM-driven Offensive Security: Testbeds, Metrics, and Experiment Design

2025-04-14 · Andreas Happe, Jürgen Cito

Large Language Models (LLMs) have emerged as a powerful approach for driving offensive penetration-testing tooling. Due to the opaque nature of LLMs, empirical methods are typically used to analyze their efficacy. The qu…

BenchmarkingLanguage ModelingLanguage ModellingLarge Language Model

Conformal Prediction for Offensive Security

2026-09-04 · Giovanni Cherubin arxiv

Despite its introduction more than a quarter century ago, Conformal Prediction (CP) has seen surprisingly few applications to the cyber security world thus far. In particular, we observe that, while CP has been employed …

Red-Teaming the Agentic Red-Team

2026-06-23 · Dario Pasquini, Michal Bazyli, Taras Fedynyshyn, Artem Sorokin arxiv

The use of agentic systems to perform offensive security operations has moved from a theoretical possibility to a commoditized capability. However, while the community has focused on creating more and more capable agents…

Building Safe GenAI Applications: An End-to-End Overview of Red Teaming for Large Language Models

2025-03-03 · Alberto Purpura, Sahil Wadhwa, Jesse Zymet, Akshay Gupta 외

The rapid growth of Large Language Models (LLMs) presents significant privacy, security, and ethical concerns. While much research has proposed methods for defending LLM systems against misuse by malicious actors, resear…

Red TeamingSurvey