paper-with-me

Papers

Near-Miss: Latent Policy Failure Detection in Agentic Workflows

2026-03-31 · Ella Rabinovich, David Boaz, Naama Zwerdling, Ateret Anaby-Tavor arxiv

Agentic systems for business process automation often require compliance with policies governing conditional updates to the system state. Evaluation of policy adherence in LLM-based agentic workflows is typically performed by comparing the final system state against a predefined ground truth. While this approach detects explicit policy violations, it may overlook a more subtle class of issues in which agents bypass required policy checks, yet reach a correct outcome due to favorable circumstances. We refer to such cases as near-misses or latent failures. In this work, we introduce a novel metric for detecting latent policy failures in agent conversations traces. Building on the ToolGuard framework, which converts natural-language policies into executable guard code, our method analyzes agent trajectories to determine whether agent's tool-calling decisions where sufficiently informed. We evaluate our approach on the $τ^2$-verified Airlines benchmark across several contemporary open and proprietary LLMs acting as agents. Our results show that latent failures occur in 8-17% of trajectories involving mutating tool calls, even when the final outcome matches the expected ground-truth state. These findings reveal a blind spot in current evaluation methodologies and highlight the need for metrics that assess not only final outcomes but also the decision process leading to them.

📄 PDF Abstract BibTeX arXiv:2603.29665

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ContactGuard: Pre-Contact Execution Monitoring with Action-Conditioned Latent World Models

2026-08-13 · Gehan Zheng, Matthew Johnson-Roberson, Weiming Zhi arxiv

Contact-rich manipulation failures are often detected only after the robot has committed to contact. This is especially limiting in wrist-camera setups: close gripper--object views help observe contact, but a poor approa…

Video Prediction

InFeR: Informed Failure Resilience in Learned Visual Navigation Control

2025-10-28 · Zishuo Wang, Joel Loo, David Hsu arxiv

While imitation learning (IL) has enabled successful visual navigation in many common environments, IL policies are prone to unpredictable failures under out-of-distribution (OOD) scenarios. This necessitates failure-res…

Visual Navigation

Resilient Legged Local Navigation: Learning to Traverse with Compromised Perception End-to-End

2023-10-05 · Jin Jin, Chong Zhang, Jonas Frey, Nikita Rudin 외

Autonomous robots must navigate reliably in unknown environments even under compromised exteroceptive perception, or perception failures. Such failures often occur when harsh environments lead to degraded sensing, or whe…

Anomaly DetectionCPUNavigateReinforcement Learning (RL)

Territory Paint Wars: Diagnosing and Mitigating Failure Modes in Competitive Multi-Agent PPO

2026-04-04 · Diyansha Singh arxiv

We present Territory Paint Wars, a minimal competitive multi-agent reinforcement learning environment implemented in Unity, and use it to systematically investigate failure modes of Proximal Policy Optimisation (PPO) und…

Multi-agent Reinforcement Learning

Latent Tensor Factorization with Nonlinear PID Control for Missing Data Recovery in Non-Intrusive Load Monitoring

2025-04-18 · Yiran Wang, Tangtang Xie, Hao Wu

Non-Intrusive Load Monitoring (NILM) has emerged as a key smart grid technology, identifying electrical device and providing detailed energy consumption data for precise demand response management. Nevertheless, NILM dat…

Computational EfficiencyMissing ValuesNon-Intrusive Load Monitoring