paper-with-me

Papers

LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures

2025-05-02 · Francisco Aguilera-Martínez, Fernando Berzal

As large language models (LLMs) continue to evolve, it is critical to assess the security threats and vulnerabilities that may arise both during their training phase and after models have been deployed. This survey seeks to define and categorize the various attacks targeting LLMs, distinguishing between those that occur during the training phase and those that affect already trained models. A thorough analysis of these attacks is presented, alongside an exploration of defense mechanisms designed to mitigate such threats. Defenses are classified into two primary categories: prevention-based and detection-based defenses. Furthermore, our survey summarizes possible attacks and their corresponding defense strategies. It also provides an evaluation of the effectiveness of the known defense mechanisms for the different security threats. Our survey aims to offer a structured framework for securing LLMs, while also identifying areas that require further research to improve and strengthen defenses against emerging security challenges.

📄 PDF Abstract BibTeX arXiv:2505.01177

Code (0)

등록된 구현이 없습니다.

Tasks

Survey

Similar Papers 제목 키워드 기반

A Survey on Agentic Security: Applications, Threats and Defenses

2025-10-07 · Asif Shahriar, Md Nafiu Rahman, Sadif Ahmed, Farig Sadeque 외 arxiv

LLM-based agents are now used throughout cybersecurity. While these agents facilitate powerful and autonomous security applications, their autonomy opens up new attack surfaces, and the security community is actively bui…

Leaky Nets: Recovering Embedded Neural Network Models and Inputs through Simple Power and Timing Side-Channels -- Attacks and Defenses

2021-03-26 · Saurav Maji, Utsav Banerjee, Anantha P. Chandrakasan

With the recent advancements in machine learning theory, many commercial embedded micro-processors use neural network models for a variety of signal processing applications. However, their associated side-channel securit…

Learning Theory

Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents

2024-10-03 · Hanrong Zhang, Jingyuan Huang, Kai Mei, Yifei Yao 외

Although LLM-based agents, powered by Large Language Models (LLMs), can use external tools and memory mechanisms to solve complex real-world tasks, they may also introduce critical security vulnerabilities. However, the …

Autonomous DrivingBackdoor AttackBenchmarking

Attacks and Defenses for Generative Diffusion Models: A Comprehensive Survey

2024-08-06 · Vu Tuan Truong, Luan Ba Dang, Long Bao Le

Diffusion models (DMs) have achieved state-of-the-art performance on various generative tasks such as image synthesis, text-to-image, and text-guided image-to-image generation. However, the more powerful the DMs, the mor…

DenoisingImage GenerationSurvey

Security of Deep Learning Methodologies: Challenges and Opportunities

2019-12-08 · Shahbaz Rezaei, Xin Liu

Despite the plethora of studies about security vulnerabilities and defenses of deep learning models, security aspects of deep learning methodologies, such as transfer learning, have been rarely studied. In this article, …

Deep LearningTransfer Learning