paper-with-me

홈 › Papers

Attack and defense techniques in large language models: A survey and new perspectives

2025-05-02 · Zhiyu Liao, Kang Chen, Yuanguo Lin, Kangkang Li, Yunxuan Liu, Hefeng Chen, Xingwang Huang, Yuanhui Yu

Large Language Models (LLMs) have become central to numerous natural language processing tasks, but their vulnerabilities present significant security and ethical challenges. This systematic survey explores the evolving landscape of attack and defense techniques in LLMs. We classify attacks into adversarial prompt attack, optimized attacks, model theft, as well as attacks on application of LLMs, detailing their mechanisms and implications. Consequently, we analyze defense strategies, including prevention-based and detection-based defense methods. Although advances have been made, challenges remain to adapt to the dynamic threat landscape, balance usability with robustness, and address resource constraints in defense implementation. We highlight open problems, including the need for adaptive scalable defenses, explainable security techniques, and standardized evaluation frameworks. This survey provides actionable insights and directions for developing secure and resilient LLMs, emphasizing the importance of interdisciplinary collaboration and ethical considerations to mitigate risks in real-world applications.

📄 PDF Abstract BibTeX arXiv:2505.00976

Code (0)

등록된 구현이 없습니다.

Tasks

Survey

Similar Papers 제목 키워드 기반

Recent advancements in LLM Red-Teaming: Techniques, Defenses, and Ethical Considerations

2024-10-09 · Tarun Raheja, Nilay Pochhi, F. D. C. M. Curie

Large Language Models (LLMs) have demonstrated remarkable capabilities in natural language processing tasks, but their vulnerability to jailbreak attacks poses significant security risks. This survey paper presents a com…

Language ModelingLanguage ModellingLarge Language ModelPrompt Engineering+1

LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures

2025-05-02 · Francisco Aguilera-Martínez, Fernando Berzal

As large language models (LLMs) continue to evolve, it is critical to assess the security threats and vulnerabilities that may arise both during their training phase and after models have been deployed. This survey seeks…

Survey

A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations

2025-02-06 · Yihe Zhou, Tao Ni, Wei-Bin Lee, Qingchuan Zhao

Large Language Models (LLMs) have achieved significantly advanced capabilities in understanding and generating human language text, which have gained increasing popularity over recent years. Apart from their state-of-the…

Safety of Multimodal Large Language Models on Images and Texts

2024-02-01 · Xin Liu, Yichen Zhu, Yunshi Lan, Chao Yang 외

Attracted by the impressive power of Multimodal Large Language Models (MLLMs), the public is increasingly utilizing them to improve the efficiency of daily work. Nonetheless, the vulnerabilities of MLLMs to unsafe instru…

Survey

Exploring Vulnerabilities and Protections in Large Language Models: A Survey

2024-06-01 · Frank Weizhen Liu, Chenhui Hu

As Large Language Models (LLMs) increasingly become key components in various AI applications, understanding their security vulnerabilities and the effectiveness of defense mechanisms is crucial. This survey examines the…

Data PoisoningSurvey