paper-with-me

홈 › Papers

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability

2026-05-09 · Hamed Omidvar, Vahideh Akhlaghi arxiv

Agents built on large language models (LLMs) rely on a range of reliability techniques, including retry, majority voting, and self-consistency, that have been developed in parallel rather than within a common analytical framework. We observe that an LLM sampled at temperature $T$ is a discrete stochastic channel $p(y \mid x)$ in the sense of Shannon's coding theory, and use this identity as the entry point for such a framework grounded in communication theory. Each of these techniques is a special case of one of six classical reliability operators: diversity combining, hybrid retransmission, iterative generator-critic decoding, rateless sampling, structured redundant verification, and difficulty-adaptive routing. Within the framework we give two closed-form results: a noise-variance threshold above which uniform averaging beats quality-weighted averaging, and a contractivity criterion for generator-critic refinement, consistent with a contractive-to-divergent transition we observe between 3B- and 14B-parameter models. We further introduce a cost-aware semantic-nearest-neighbor router whose single Lagrangian knob traverses the quality-cost frontier without retraining. Across six channel configurations spanning local and cloud models on 69 hard tasks, no fixed model-technique-budget choice dominates, motivating per-task allocation. On a 300-item hard split of MMLU, GSM8K, and HumanEval, our router occupies the full empirical Pareto frontier: at matched quality, its normalized cost is ${\approx}56$\% lower than the strongest fixed technique; at matched normalized cost, it improves quality by ${\approx}7$\% ($26$\% over single-shot decoding). These results argue for consolidating these reliability techniques into a single tunable layer informed by channel coding.

📄 PDF Abstract BibTeX arXiv:2605.09121

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Communication to Completion: Modeling Collaborative Workflows with Intelligent Multi-Agent Communication

2025-10-22 · Yiming Lu, Xun Wang, Simin Ma, Shujian Liu 외 arxiv

Multi-agent LLM systems have demonstrated impressive capabilities in complex collaborative tasks, yet most frameworks treat communication as instantaneous and free, overlooking a fundamental constraint in real world team…

SkillFlow: Efficient Skill and Code Transfer Through Communication in Adapting AI Agents

2025-04-08 · Pagkratios Tagkopoulos, Fangzhou Li, Ilias Tagkopoulos

AI agents are autonomous systems that can execute specific tasks based on predefined programming. Here, we present SkillFlow, a modular, technology-agnostic framework that allows agents to expand their functionality in a…

Scheduling

When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding

2026-05-31 · Xu Yang, Lunyiu Nie, Ethan Chandra, Stanislav Gannutin 외 arxiv

Multi-agent Large Language Model (LLM) systems offer a way to decompose complex tasks, such as coding, through parallelization and context isolation. However, adding agents in practice introduces inter-agent communicatio…

Community Detectiongraph partitioning

Privacy-Preserving Communication-Efficient Federated Multi-Armed Bandits

2021-11-02 · Tan Li, Linqi Song

Communication bottleneck and data privacy are two critical concerns in federated multi-armed bandit (MAB) problems, such as situations in decision-making and recommendations of connected vehicles via wireless. In this pa…

Decision MakingMulti-Armed BanditsPrivacy Preserving

Federated Contextual Cascading Bandits with Asynchronous Communication and Heterogeneous Users

2024-02-26 · Hantao Yang, Xutong Liu, Zhiyong Wang, Hong Xie 외

We study the problem of federated contextual combinatorial cascading bandits, where $|\mathcal{U}|$ agents collaborate under the coordination of a central server to provide tailored recommendations to the $|\mathcal{U}|$…