paper-with-me

홈 › Papers

CORBA: Contagious Recursive Blocking Attacks on Multi-Agent Systems Based on Large Language Models

2025-02-20 · Zhenhong Zhou, Zherui Li, Jie Zhang, Yuanhe Zhang, Kun Wang, Yang Liu, Qing Guo

Large Language Model-based Multi-Agent Systems (LLM-MASs) have demonstrated remarkable real-world capabilities, effectively collaborating to complete complex tasks. While these systems are designed with safety mechanisms, such as rejecting harmful instructions through alignment, their security remains largely unexplored. This gap leaves LLM-MASs vulnerable to targeted disruptions. In this paper, we introduce Contagious Recursive Blocking Attacks (Corba), a novel and simple yet highly effective attack that disrupts interactions between agents within an LLM-MAS. Corba leverages two key properties: its contagious nature allows it to propagate across arbitrary network topologies, while its recursive property enables sustained depletion of computational resources. Notably, these blocking attacks often involve seemingly benign instructions, making them particularly challenging to mitigate using conventional alignment methods. We evaluate Corba on two widely-used LLM-MASs, namely, AutoGen and Camel across various topologies and commercial models. Additionally, we conduct more extensive experiments in open-ended interactive LLM-MASs, demonstrating the effectiveness of Corba in complex topology structures and open-source models. Our code is available at: https://github.com/zhrli324/Corba.

📄 PDF Abstract BibTeX arXiv:2502.14529

Code (1)

zhrli324/corba 공식 구현

Tasks

BlockingLanguage ModelingLanguage ModellingLarge Language Model

Similar Papers 제목 키워드 기반

Optimal Investment under Information Driven Contagious Distress

2016-12-19

We introduce a dynamic optimization framework to analyze optimal portfolio allocations within an information driven contagious distress model. The investor allocates his wealth across several stocks whose growth rates an…

A Troublemaker with Contagious Jailbreak Makes Chaos in Honest Towns

2024-10-21 · Tianyi Men, Pengfei Cao, Zhuoran Jin, Yubo Chen 외

With the development of large language models, they are widely used as agents in various fields. A key component of agents is memory, which stores vital information but is susceptible to jailbreak attacks. Existing resea…

Attribute

Rethinking Latency Denial-of-Service: Attacking the LLM Serving Framework, Not the Model

2026-02-08 · Tianyi Wang, Huawei Fan, Yuanchao Shu, Peng Cheng 외 arxiv

Large Language Models face an emerging and critical threat known as latency attacks. Because LLM inference is inherently expensive, even modest slowdowns can translate into substantial operating costs and severe availabi…

Prompt Engineering

AdVersarial: Perceptual Ad Blocking meets Adversarial Machine Learning

2018-11-08 · Florian Tramèr, Pascal Dupré, Gili Rusak, Giancarlo Pellegrino 외

Perceptual ad-blocking is a novel approach that detects online advertisements based on their visual content. Compared to traditional filter lists, the use of perceptual signals is believed to be less prone to an arms rac…

BIG-bench Machine LearningBlocking

Learning to Discriminate Perturbations for Blocking Adversarial Attacks in Text Classification

2019-09-06 · IJCNLP 2019 11 · Yichao Zhou, Jyun-Yu Jiang, Kai-Wei Chang, Wei Wang

Adversarial attacks against machine learning models have threatened various real-world applications such as spam filtering and sentiment analysis. In this paper, we propose a novel framework, learning to DIScriminate Per…

BlockingGeneral ClassificationSentiment Analysistext-classification+1