paper-with-me

홈 › Papers

Securing Large Language Models: Threats, Vulnerabilities and Responsible Practices

2024-03-19 · Sara Abdali, Richard Anarfi, CJ Barberan, Jia He, Erfan Shayegani

Large language models (LLMs) have significantly transformed the landscape of Natural Language Processing (NLP). Their impact extends across a diverse spectrum of tasks, revolutionizing how we approach language understanding and generations. Nevertheless, alongside their remarkable utility, LLMs introduce critical security and risk considerations. These challenges warrant careful examination to ensure responsible deployment and safeguard against potential vulnerabilities. This research paper thoroughly investigates security and privacy concerns related to LLMs from five thematic perspectives: security and privacy concerns, vulnerabilities against adversarial attacks, potential harms caused by misuses of LLMs, mitigation strategies to address these challenges while identifying limitations of current strategies. Lastly, the paper recommends promising avenues for future research to enhance the security and risk management of LLMs.

📄 PDF Abstract BibTeX arXiv:2403.12503

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures

2025-05-02 · Francisco Aguilera-Martínez, Fernando Berzal

As large language models (LLMs) continue to evolve, it is critical to assess the security threats and vulnerabilities that may arise both during their training phase and after models have been deployed. This survey seeks…

Survey

MCP-Guard: A Multi-Stage Defense-in-Depth Framework for Securing Model Context Protocol in Agentic AI

2025-08-14 · Wenpeng Xing, Zhonghao Qi, Yupeng Qin, Yilin Li 외 arxiv

While Large Language Models (LLMs) have achieved remarkable performance, they remain vulnerable to jailbreak. The integration of Large Language Models (LLMs) with external tools via protocols such as the Model Context Pr…

"Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills in the Wild

2026-02-06 · Yi Liu, Zhihao Chen, Yanjun Zhang, Gelei Deng 외 arxiv

LLM-based coding agents increasingly rely on third-party extensions called skills, which bundle natural language instructions and helper scripts that execute with full user privileges. Community registries have emerged t…

A Survey on Adversarial Robustness of LiDAR-based Machine Learning Perception in Autonomous Vehicles

2024-11-21 · Junae Kim, Amardeep Kaur

In autonomous driving, the combination of AI and vehicular technology offers great potential. However, this amalgamation comes with vulnerabilities to adversarial attacks. This survey focuses on the intersection of Adver…

Adversarial RobustnessAutonomous DrivingAutonomous Vehicles

Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025

2025-06-14 · Zonghao Ying, Siyang Wu, Run Hao, Peng Ying 외

Multimodal Large Language Models (MLLMs) have enabled transformative advancements across diverse applications but remain susceptible to safety threats, especially jailbreak attacks that induce harmful outputs. To systema…