paper-with-me

Papers

MTTR-A: Measuring Cognitive Recovery Latency in Multi-Agent Systems

2025-11-08 · Barak Or arxiv

Reliability in multi-agent systems (MAS) built on large language models is increasingly limited by cognitive failures rather than infrastructure faults. Existing observability tools describe failures but do not quantify how quickly distributed reasoning recovers once coherence is lost. We introduce MTTR-A (Mean Time-to-Recovery for Agentic Systems), a runtime reliability metric that measures cognitive recovery latency in MAS. MTTR-A adapts classical dependability theory to agentic orchestration, capturing the time required to detect reasoning drift and restore coherent operation. We further define complementary metrics, including MTBF and a normalized recovery ratio (NRR), and establish theoretical bounds linking recovery latency to long-run cognitive uptime. Using a LangGraph-based benchmark with simulated drift and reflex recovery, we empirically demonstrate measurable recovery behavior across multiple reflex strategies. This work establishes a quantitative foundation for runtime cognitive dependability in distributed agentic systems.

📄 PDF Abstract BibTeX arXiv:2511.20663

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Resiliency in Large Language Model Serving with KevlarFlow

2026-01-30 · Shangshu Qian, Kipling Liu, P. C. Sruthi, Lin Tan 외 arxiv

Large Language Model (LLM) serving systems remain fundamentally fragile, where frequent hardware faults in hyperscale clusters trigger disproportionate service outages in the software stack. Current recovery mechanisms a…

From Automated to Autonomous: Hierarchical Agent-native Network Architecture (HANA)

2026-05-20 · Binghan Wu, Shoufeng Wang, Yunxin Liu, Ya-Qin Zhang 외 arxiv

Realizing Level 4/5 Autonomous Networks (AN) demands a shift from static automation to agent-native intelligence. Current operations, reliant on rigid scripts, lack the cognitive agency to handle off-nominal conditions. …

End-to-End Referring Video Object Segmentation with Multimodal Transformers

2021-11-29 · CVPR 2022 1 · Adam Botach, Evgenii Zheltonozhskii, Chaim Baskin

The referring video object segmentation task (RVOS) involves segmentation of a text-referred object instance in the frames of a given video. Due to the complex nature of this multimodal task, which combines text reasonin…

Inductive BiasInstance SegmentationReferring Expression SegmentationReferring Video Object Segmentation+5

Agentic Observability: Automated Alert Triage for Adobe E-Commerce

2026-01-31 · Aprameya Bharadwaj, Kyle Tu arxiv

Modern enterprise systems exhibit complex interdependencies that make observability and incident response increasingly challenging. Manual alert triage, which typically involves log inspection, API verification, and cros…

Prioritizing Security Practice Adoption: Empirical Insights on Software Security Outcomes in the npm Ecosystem

2025-04-18 · Nusrat Zahan, Laurie Williams

Practitioners often struggle with the overwhelming number of security practices outlined in cybersecurity frameworks for risk mitigation. Given the limited budget, time, and resources, practitioners want to prioritize th…