paper-with-me

Papers

GRADE: Graph Representation of LLM Agent Dependency and Execution

2026-06-22 · Yue Zhao arxiv

Can one graph represent every kind of LLM agent's run? A trace records what each step did, never what it relied on, the state it read, and the results it reused. GRADE recovers that missing layer: it models any run as one graph over its step nodes with two edge layers, execution edges (what ran in what order) read from the trace for free, and dependency edges (what each step relied on) rarely logged, so each is graded by how it is known, observed, declared, or inferred. One representation, and each layer earns its place. Across six corpora of LLM agents spanning tool use, coding, and the web, the dependency layer can predict failure where run size is weak and, under leave-one-corpus-out transfer, stays above chance on every held-out class while run size fails. Meanwhile, the execution layer localizes the faulting step in a failed multi-agent run. This work also provides a more in-depth analysis of why generic graph neural networks may misread the dependency layer, unlike our feature-based alternative. The same graph representation opens further uses, carrying from failure diagnosis in a single run to efficiency and robustness optimization at scale.

📄 PDF Abstract BibTeX arXiv:2606.22741

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams

2026-05-25 · Tianda Sun, Dimitar Kazakov arxiv

Tool-using LLM agents produce trajectories whose calls form a directed dependency graph: earlier tool outputs supply arguments to later calls. Whether this execution structure is represented inside the model is unknown; …

From Sequence to Structure: Relational Uncertainty Propagation for LLM Agents

2026-08-17 · Zhengzhao Ma. Boxi Cao, Yaojie Lu, Hongyu Lin, Xianpei Han 외 hf

Reliable uncertainty quantification (UQ) is essential for deploying large language model (LLM) agents in complex interactive environments. Existing UQ methods largely rely on local signals, such as token probabilities, p…

The Six Sigma Agent: Achieving Enterprise-Grade Reliability in LLM Systems Through Consensus-Driven Decomposed Execution

2026-01-29 · Khush Patel, Siva Surendira, Jithin George, Shreyas Kapale arxiv

Large Language Models demonstrate remarkable capabilities yet remain fundamentally probabilistic, presenting critical reliability challenges for enterprise deployment. We introduce the Six Sigma Agent, a novel architectu…

A Feedback Scheme to Reorder a Multi-Agent Execution Schedule by Persistently Optimizing a Switchable Action Dependency Graph

2020-10-11 · Alexander Berndt, Niels van Duijkeren, Luigi Palmieri, Tamas Keviczky

In this paper we consider multiple Automated Guided Vehicles (AGVs) navigating a common workspace to fulfill various intralogistics tasks, typically formulated as the Multi-Agent Path Finding (MAPF) problem. To keep plan…

ManagementMulti-Agent Path Finding

GAP: Graph-Based Agent Planning with Parallel Tool Use and Reinforcement Learning

2025-10-29 · Jiaqi Wu, Qinlao Zhao, Zefeng Chen, Kai Qin 외 arxiv

Autonomous agents powered by large language models (LLMs) have shown impressive capabilities in tool manipulation for complex task-solving. However, existing paradigms such as ReAct rely on sequential reasoning and execu…

Multi-hop Question AnsweringReinforcement Learning