paper-with-me

홈 › Papers

Security Assessment and Mitigation Strategies for Large Language Models: A Comprehensive Defensive Framework

2026-03-17 · Taiwo Onitiju, Iman Vakilinia arxiv

Large Language Models increasingly power critical infrastructure from healthcare to finance, yet their vulnerability to adversarial manipulation threatens system integrity and user safety. Despite growing deployment, no comprehensive comparative security assessment exists across major LLM architectures, leaving organizations unable to quantify risk or select appropriately secure LLMs for sensitive applications. This research addresses this gap by establishing a standardized vulnerability assessment framework and developing a multi-layered defensive system to protect against identified threats. We systematically evaluate five widely-deployed LLM families GPT-4, GPT-3.5 Turbo, Claude-3 Haiku, LLaMA-2-70B, and Gemini-2.5-pro against 10,000 adversarial prompts spanning six attack categories. Our assessment reveals critical security disparities, with vulnerability rates ranging from 11.9\% to 29.8\%, demonstrating that LLM capability does not correlate with security robustness. To mitigate these risks, we develop a production-ready defensive framework achieving 83\% average detection accuracy with only 5\% false positives. These results demonstrate that systematic security assessment combined with external defensive measures provides a viable path toward safer LLM deployment in production environments.

📄 PDF Abstract BibTeX arXiv:2603.17123

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems

2024-01-11 · Tianyu Cui, Yanling Wang, Chuanpu Fu, Yong Xiao 외

Large language models (LLMs) have strong capabilities in solving diverse natural language processing tasks. However, the safety and security issues of LLM systems have become the major obstacle to their widespread applic…

Language ModelingLanguage ModellingLarge Language Model

Large Language Models in Cybersecurity: Applications, Vulnerabilities, and Defense Techniques

2025-07-18 · Niveen O. Jaffal, Mohammed Alkhanafseh, David Mohaisen arxiv

Large Language Models (LLMs) are transforming cybersecurity by enabling intelligent, adaptive, and automated approaches to threat detection, vulnerability assessment, and incident response. With their advanced language u…

A Formal Framework for Assessing and Mitigating Emergent Security Risks in Generative AI Models: Bridging Theory and Dynamic Risk Mitigation

2024-10-15 · Aviral Srivastava, Sourav Panda

As generative AI systems, including large language models (LLMs) and diffusion models, advance rapidly, their growing adoption has led to new and complex security risks often overlooked in traditional AI risk assessment …

Anomaly DetectionRed Teaming

Enterprise-Grade Security for the Model Context Protocol (MCP): Frameworks and Mitigation Strategies

2025-04-11 · Vineeth Sai Narajala, Idan Habler

The Model Context Protocol (MCP), introduced by Anthropic, provides a standardized framework for artificial intelligence (AI) systems to interact with external data sources and tools in real-time. While MCP offers signif…

Towards Automated Network Mitigation Analysis (extended)

2017-05-15 · Patrick Speicher, Marcel Steinmetz, Jörg Hoffmann, Michael Backes 외

Penetration testing is a well-established practical concept for the identification of potentially exploitable security weaknesses and an important component of a security audit. Providing a holistic security assessment f…