paper-with-me

홈 › Papers

Rethinking the Reliability of Multi-agent System: A Perspective from Byzantine Fault Tolerance

2025-11-13 · Lifan Zheng, Jiawei Chen, Qinghong Yin, Jingyuan Zhang, Xinyi Zeng, Yu Tian arxiv

Ensuring the reliability of agent architectures and effectively identifying problematic agents when failures occur are crucial challenges in multi-agent systems (MAS). Advances in large language models (LLMs) have established LLM-based agents as a major branch of MAS, enabling major breakthroughs in complex problem solving and world modeling. However, the reliability implications of this shift remain largely unexplored. i.e., whether substituting traditional agents with LLM-based agents can effectively enhance the reliability of MAS. In this work, we investigate and quantify the reliability of LLM-based agents from the perspective of Byzantine fault tolerance. We observe that LLM-based agents demonstrate stronger skepticism when processing erroneous message flows, a characteristic that enables them to outperform traditional agents across different topological structures. Motivated by the results of the pilot experiment, we design CP-WBFT, a confidence probe-based weighted Byzantine Fault Tolerant consensus mechanism to enhance the stability of MAS with different topologies. It capitalizes on the intrinsic reflective and discriminative capabilities of LLMs by employing a probe-based, weighted information flow transmission method to improve the reliability of LLM-based agents. Extensive experiments demonstrate that CP-WBFT achieves superior performance across diverse network topologies under extreme Byzantine conditions (85.7\% fault rate). Notably, our approach surpasses traditional methods by attaining remarkable accuracy on various topologies and maintaining strong reliability in both mathematical reasoning and safety assessment tasks.

📄 PDF Abstract BibTeX arXiv:2511.10400

Code (0)

등록된 구현이 없습니다.

Tasks

Mathematical Reasoning

Similar Papers 제목 키워드 기반

Rethinking Failure Attribution in Multi-Agent Systems: A Multi-Perspective Benchmark and Evaluation

2026-03-26 · Yeonjun In, Mehrab Tanjim, Jayakumar Subramanian, Sungchul Kim 외 arxiv

Failure attribution is essential for diagnosing and improving multi-agent systems (MAS), yet existing benchmarks and methods largely assume a single deterministic root cause for each failure. In practice, MAS failures of…

Rethinking Noisy Video-Text Retrieval via Relation-aware Alignment

2025-01-01 · CVPR 2025 1 · Huakai Lai, Guoxin Xiong, Huayu Mai, Xiang Liu 외

Video-Text Retrieval (VTR) is a core task in multi-modal understanding, drawing growing attention from both academia and industry in recent years. While numerous VTR methods have achieved success, most of them assume…

RelationRetrievalText RetrievalVideo-Text Retrieval

Looking Forward: Challenges and Opportunities in Agentic AI Reliability

2025-11-14 · Liudong Xing, Janet, Lin arxiv

This chapter presents perspectives for challenges and future development in building reliable AI systems, particularly, agentic AI systems. Several open research problems related to mitigating the risks of cascading fail…

Allen: Rethinking MAS Design through Step-Level Policy Autonomy

2025-08-15 · Qiangong Zhou, Zhiting Wang, Mingyou Yao, Zongyang Liu arxiv

We introduce a new Multi-Agent System (MAS) - Allen, designed to address two core challenges in current MAS design: (1) improve system's policy autonomy, empowering agents to dynamically adapt their behavioral strategies…

Perspectives on a Reliability Monitoring Framework for Agentic AI Systems

2025-11-12 · Niclas Flehmig, Mary Ann Lundteigen, Shen Yin arxiv

The implementation of agentic AI systems has the potential of providing more helpful AI systems in a variety of applications. These systems work autonomously towards a defined goal with reduced external control. Despite …

Out-of-Distribution Detection