paper-with-me

홈 › Papers

Every Response Counts: Quantifying Uncertainty of LLM-based Multi-Agent Systems through Tensor Decomposition

2026-04-09 · Tiejin Chen, Huaiyuan Yao, Jia Chen, Evangelos E. Papalexakis, Hua Wei arxiv

While Large Language Model-based Multi-Agent Systems (MAS) consistently outperform single-agent systems on complex tasks, their intricate interactions introduce critical reliability challenges arising from communication dynamics and role dependencies. Existing Uncertainty Quantification methods, typically designed for single-turn outputs, fail to address the unique complexities of the MAS. Specifically, these methods struggle with three distinct challenges: the cascading uncertainty in multi-step reasoning, the variability of inter-agent communication paths, and the diversity of communication topologies. To bridge this gap, we introduce MATU, a novel framework that quantifies uncertainty through tensor decomposition. MATU moves beyond analyzing final text outputs by representing entire reasoning trajectories as embedding matrices and organizing multiple execution runs into a higher-order tensor. By applying tensor decomposition, we disentangle and quantify distinct sources of uncertainty, offering a comprehensive reliability measure that is generalizable across different agent structures. We provide comprehensive experiments to show that MATU effectively estimates holistic and robust uncertainty across diverse tasks and communication topologies.

📄 PDF Abstract BibTeX arXiv:2604.08708

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Conformal Prediction Sets Improve Human Decision Making

2024-01-24 · Jesse C. Cresswell, Yi Sui, Bhargava Kumar, Noël Vouitsis

In response to everyday queries, humans explicitly signal uncertainty and offer alternative answers when they are unsure. Machine learning models that output calibrated prediction sets through conformal prediction mimic …

Conformal PredictionDecision MakingPrediction

Semantic Volume: Quantifying and Detecting both External and Internal Uncertainty in LLMs

2025-02-28 · Xiaomin Li, Zhou Yu, Ziji Zhang, Yingying Zhuang 외

Large language models (LLMs) have demonstrated remarkable performance across diverse tasks by encoding vast amounts of factual knowledge. However, they are still prone to hallucinations, generating incorrect or misleadin…

Hallucination

DiverseAgentEntropy: Quantifying Black-Box LLM Uncertainty through Diverse Perspectives and Multi-Agent Interaction

2024-12-12 · Yu Feng, Phu Mon Htut, Zheng Qi, Wei Xiao 외

Quantifying the uncertainty in the factual parametric knowledge of Large Language Models (LLMs), especially in a black-box setting, poses a significant challenge. Existing methods, which gauge a model's uncertainty throu…

Can Multiple Responses from an LLM Reveal the Sources of Its Uncertainty?

2025-08-28 · Yang Nan, Pengfei He, Ravi Tandon, Han Xu arxiv

Large language models (LLMs) have delivered significant breakthroughs across diverse domains but can still produce unreliable or misleading outputs, posing critical challenges for real-world applications. While many rece…

Quantifying Uncertainty in Answers from any Language Model and Enhancing their Trustworthiness

2023-08-30 · Jiuhai Chen, Jonas Mueller

We introduce BSDetector, a method for detecting bad and speculative answers from a pretrained Large Language Model by estimating a numeric confidence score for any output it generated. Our uncertainty quantification tech…

Language ModelingLanguage ModellingLarge Language ModelUncertainty Quantification