paper-with-me

Papers

Formal Verification of Agentic Systems over Operational Data

2026-08-04 · Alejandro J. Mercado, Alessio Lomuscio arxiv

Agentic systems driven by large language models (LLMs) are increasingly deployed in real-world workflows where they act on persistent operational data. Before deployment, these systems need to be verified against business requirements that govern workflow execution and data evolution. However, existing approaches do not provide such system-level guarantees, as they mainly constrain or analyse behaviour at the agent's interface level. We study here the verification of agentic systems comprising a single LLM and a tool orchestration harness over relational operational data. We formalise them as Stateful Tool-Enabled Agentic Deployments (STEADs), give their semantics, define the problem of verifying them against First-Order Computation Tree Logic (FO-CTL) specifications, and show that it is undecidable. We identify sufficient conditions for exact preservation of FO-CTL specifications under a finite-domain restriction, over which verification is PSPACE-complete. The key requirement is that renaming opaque identifiers in the data must correspondingly rename the selected tool calls. We show that LLM-driven agents can violate this condition and introduce a canonical deployment wrapper that guarantees it for arbitrary base agents while preserving already-equivariant behaviour. We prove that computing canonical representations required by this construction is graph-isomorphism-hard. Finally, we illustrate our framework on an LLM agent orchestrating a case-management workflow.

📄 PDF Abstract BibTeX arXiv:2608.03609

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Architecting Agentic Communities using Design Patterns

2026-01-07 · Zoran Milosevic, Fethi Rabhi arxiv

The rapid evolution of Large Language Models (LLM) and subsequent Agentic AI technologies requires systematic architectural guidance for building sophisticated, production-grade systems. This paper presents an approach f…

Containment Verification: AI Safety Guarantees Independent of Alignment

2026-05-09 · Royce Moon, Lav R. Varshney arxiv

Agentic frameworks are the software layer through which AI agents act in the world. Existing safety methods intervene on the model and therefore remain conditional on unverifiable properties of learned behavior. We intro…

Aether: Network Validation Using Agentic AI and Digital Twin

2026-04-20 · Jordan Auge, Sam Betts, Giovanna Carofiglio, Giulio Grassi 외 arxiv

Network change validation remains a critical yet predominantly manual, time-consuming, and error-prone process in modern network operations. While formal network verification has made substantial progress in proving corr…

BarrierBench: Evaluating Large Language Models for Safety Verification in Dynamical Systems

2025-11-12 · Ali Taheri, Alireza Taban, Sadegh Soudjani, Ashutosh Trivedi arxiv

Safety verification of dynamical systems via barrier certificates is essential for ensuring correctness in autonomous applications. Synthesizing these certificates involves discovering mathematical functions with current…

Agentic AI-based Coverage Closure for Formal Verification

2026-03-03 · Sivaram Pothireddypalli, Ashish Raman, Deepak Narayan Gadde, Aman Kumar arxiv

Coverage closure is a critical requirement in Integrated Chip (IC) development process and key metric for verification sign-off. However, traditional exhaustive approaches often fail to achieve full coverage within proje…