paper-with-me

홈 › Papers

Agentic Problem Frames: A Systematic Approach to Engineering Reliable Domain Agents

2026-02-22 · Chanjin Park arxiv

Large Language Models (LLMs) are evolving into autonomous agents, yet current "frameless" development--relying on ambiguous natural language without engineering blueprints--leads to critical risks such as scope creep and open-loop failures. To ensure industrial-grade reliability, this study proposes Agentic Problem Frames (APF), a systematic engineering framework that shifts focus from internal model intelligence to the structured interaction between the agent and its environment. The APF establishes a dynamic specification paradigm where intent is concretized at runtime through domain knowledge injection. At its core, the Act-Verify-Refine (AVR) loop functions as a closed-loop control system that transforms execution results into verified knowledge assets, driving system behavior toward asymptotic convergence to mission requirements (R). To operationalize this, this study introduces the Agentic Job Description (AJD), a formal specification tool that defines jurisdictional boundaries, operational contexts, and epistemic evaluation criteria. The efficacy of this framework is validated through two contrasting case studies: a delegated proxy model for business travel and an autonomous supervisor model for industrial equipment management. By applying AJD-based specification and APF modeling to these scenarios, the analysis demonstrates how operational scenarios are systematically controlled within defined boundaries. These cases provide a conceptual proof that agent reliability stems not from a model's internal reasoning alone, but from the rigorous engineering structures that anchor stochastic AI within deterministic business processes, thereby enabling the development of verifiable and dependable domain agents.

📄 PDF Abstract BibTeX arXiv:2602.19065

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Agentic Large Language Models for Automated Structural Analysis of 3D Frame Systems

2026-06-02 · Ziheng Geng, Ian Franklin, Santiago Martinez, Jiachen Liu 외 arxiv

Large language models (LLMs) have emerged as powerful foundation models with strong reasoning capabilities across domains. Beyond reactive text generation, agentic LLMs enable autonomous workflow execution through modula…

Code TranslationText Generation

Agentic DraCor and the Art of Docstring Engineering: Evaluating MCP-empowered LLM Usage of the DraCor API

2025-08-19 · Peer Trilcke, Ingo Börner, Henny Sluyter-Gäthje, Daniil Skorinkin 외 arxiv

This paper reports on the implementation and evaluation of a Model Context Protocol (MCP) server for DraCor, enabling Large Language Models (LLM) to autonomously interact with the DraCor API. We conducted experiments foc…

Evaluating Verified Autonomy in Quantum Engineering

2026-09-15 · Naixu Guo, Changhao Li, Siyu Cheng, Qicheng Tang 외 arxiv

Reliable quantum engineering is essential for turning quantum phenomena into practical technologies. As quantum platforms grow in scale and complexity, their characterization and operation require increasing human effort…

From Goals to Aspects, Revisited: An NFR Pattern Language for Agentic AI Systems

2026-02-28 · Yijun Yu arxiv

Agentic AI systems exhibit numerous crosscutting concerns -- security, observability, cost management, fault tolerance -- that are poorly modularized in current implementations, contributing to the high failure rate of A…

Beyond single-channel agentic benchmarking

2026-02-05 · Nelu D. Radpour arxiv

Contemporary benchmarks for agentic artificial intelligence (AI) frequently evaluate safety through isolated task-level accuracy thresholds, implicitly treating autonomous systems as single points of failure. This single…