paper-with-me

Papers

Multi-Agent Framework for Threat Mitigation and Resilience in AI-Based Systems

2025-12-29 · Armstrong Foundjem, Lionel Nganyewou Tidjon, Leuson Da Silva, Foutse Khomh arxiv

Machine learning (ML) underpins foundation models in finance, healthcare, and critical infrastructure, making them targets for data poisoning, model extraction, prompt injection, automated jailbreaking, and preference-guided black-box attacks that exploit model comparisons. Larger models can be more vulnerable to introspection-driven jailbreaks and cross-modal manipulation. Traditional cybersecurity lacks ML-specific threat modeling for foundation, multimodal, and RAG systems. Objective: Characterize ML security risks by identifying dominant TTPs, vulnerabilities, and targeted lifecycle stages. Methods: We extract 93 threats from MITRE ATLAS (26), AI Incident Database (12), and literature (55), and analyze 854 GitHub/Python repositories. A multi-agent RAG system (ChatGPT-4o, temp 0.4) mines 300+ articles to build an ontology-driven threat graph linking TTPs, vulnerabilities, and stages. Results: We identify unreported threats including commercial LLM API model stealing, parameter memorization leakage, and preference-guided text-only jailbreaks. Dominant TTPs include MASTERKEY-style jailbreaking, federated poisoning, diffusion backdoors, and preference optimization leakage, mainly impacting pre-training and inference. Graph analysis reveals dense vulnerability clusters in libraries with poor patch propagation. Conclusion: Adaptive, ML-specific security frameworks, combining dependency hygiene, threat intelligence, and monitoring, are essential to mitigate supply-chain and inference risks across the ML lifecycle.

📄 PDF Abstract BibTeX arXiv:2512.23132

Code (0)

등록된 구현이 없습니다.

Tasks

Model extraction

Similar Papers 제목 키워드 기반

QSAF: A Novel Mitigation Framework for Cognitive Degradation in Agentic AI

2025-07-21 · Hammad Atta, Muhammad Zeeshan Baig, Yasir Mehmood, Nadeem Shahzad 외 arxiv

We introduce Cognitive Degradation as a novel vulnerability class in agentic AI systems. Unlike traditional adversarial external threats such as prompt injection, these failures originate internally, arising from memory …

RobResilience: Implementing and Evaluating a Resilience Framework for Cyber-Physical Embodied Systems

2026-09-15 · Gysella Imrell, Emanuele Miotto, Mahya Mohammadi Kashani, Mauro Conti 외 arxiv

In embodied cyber-physical systems, active cyberattacks pose an immediate threat not just to data, but to physical integrity and human safety. While existing security approaches excel at detection, they lack the runtime …

Autonomous AI-based Cybersecurity Framework for Critical Infrastructure: Real-Time Threat Mitigation

2025-07-10 · Jenifer Paulraj, Brindha Raghuraman, Nagarani Gopalakrishnan, Yazan Otoum arxiv

Critical infrastructure systems, including energy grids, healthcare facilities, transportation networks, and water distribution systems, are pivotal to societal stability and economic resilience. However, the increasing …

Vulnerability Detection

Robustifying 3D Perception through Least-Squares Multi-Agent Graphs Object Tracking

2025-07-07 · Maria Damanaki, Ioulia Kapsali, Nikos Piperigkos, Alexandros Gkillas 외

The critical perception capabilities of EdgeAI systems, such as autonomous vehicles, are required to be resilient against adversarial threats, by enabling accurate identification and localization of multiple objects in t…

Autonomous VehiclesObject Tracking

Foundations of Cyber Resilience: The Confluence of Game, Control, and Learning Theories

2024-04-01 · Quanyan Zhu

Cyber resilience is a complementary concept to cybersecurity, focusing on the preparation, response, and recovery from cyber threats that are challenging to prevent. Organizations increasingly face such threats in an evo…

Meta-Learning