paper-with-me

홈 › Papers

A Pattern Language for Resilient Visual Agents

2026-04-30 · Habtom Kahsay Gidey, Alexander Lenz, Alois Knoll arxiv

Integrating multimodal foundation models into enterprise ecosystems presents a fundamental software architecture challenge. Architects must balance competing quality attributes: the high latency and non-determinism of vision language action (VLA) models versus the strict determinism and real-time performance required by enterprise control loops. In this study, we propose an architectural pattern language for visual agents that separates fast, deterministic reflexes from slow, probabilistic supervision. It consists of four architectural design patterns: (1) Hybrid Affordance Integration, (2) Adaptive Visual Anchoring, (3) Visual Hierarchy Synthesis, and (4) Semantic Scene Graph.

📄 PDF Abstract BibTeX arXiv:2604.28001

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Resilient Consensus in Agentic AI

2026-06-12 · Sribalaji C. Anand, George J. Pappas arxiv

Large language model (LLM) agents are increasingly deployed in multi-agent systems where they must coordinate and agree on shared decisions. We ask whether classical resilient consensus theory, developed for deterministi…

Architecting Resilient LLM Agents: A Guide to Secure Plan-then-Execute Implementations

2025-09-10 · Ron F. Del Rosario, Klaudia Krawiecka, Christian Schroeder de Witt arxiv

As Large Language Model (LLM) agents become increasingly capable of automating complex, multi-step tasks, the need for robust, secure, and predictable architectural patterns is paramount. This paper provides a comprehens…

Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight

2025-09-12 · Jingyu Tang, Chaoran Chen, Jiawen Li, Zhiping Zhang 외 arxiv

The dark patterns, deceptive interface designs manipulating user behaviors, have been extensively studied for their effects on human decision-making and autonomy. Yet, with the rising prominence of LLM-powered GUI agents…

Conflict-Resilient Multi-Agent Reasoning via Signed Graph Modeling

2026-05-19 · Longgang He, Longzhu He, Daojing He, Chaozhuo Li arxiv

LLM-based multi-agent systems (MAS) have demonstrated strong reasoning and decision-making capabilities that consistently surpass those of single LLM agents. However, their performance often suffers from naive aggregatio…

See and Remember: A Multimodal Agent for Web Traversal

2026-03-03 · Xinjun Wang, Shengyao Wang, Aimin Zhou, Hao Hao arxiv

Autonomous web navigation requires agents to perceive complex visual environments and maintain long-term context, yet current Large Language Model (LLM) based agents often struggle with spatial disorientation and navigat…

Visual Grounding