paper-with-me

홈 › Papers

PowerChain: A Verifiable Agentic AI System for Automating Distribution Grid Analyses

2025-08-23 · Emmanuel O. Badmus, Peng Sang, Dimitrios Stamoulis, Amritanshu Pandey arxiv

Rapid electrification and decarbonization are increasing the complexity of distribution grid (DG) operation and planning, necessitating advanced computational analyses to ensure reliability and resilience. These analyses depend on disparate workflows comprising complex models, function calls, and data pipelines that require substantial expert knowledge and remain difficult to automate. Workforce and budget constraints further limit utilities' ability to apply such analyses at scale. To address this gap, we build an agentic system PowerChain, which is capable of autonomously performing complex grid analyses. Existing agentic AI systems are typically developed in a bottom-up manner with customized context for predefined analysis tasks; therefore, they do not generalize to tasks that the agent has never seen. In comparison, to generalize to unseen DG analysis tasks, PowerChain dynamically generates structured context by leveraging supervisory signals from self-contained power systems tools (e.g., GridLAB-D) and an optimized set of expert-annotated and verified reasoning trajectories. For complex DG tasks defined in natural language, empirical results on real utility data demonstrate that PowerChain achieves up to a 144/% improvement in performance over baselines.

📄 PDF Abstract BibTeX arXiv:2508.17094

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AgenticShop: Benchmarking Agentic Product Curation for Personalized Web Shopping

2026-02-12 · Sunghwan Kim, Ryang Heo, Yongsik Seo, Jinyoung Yeo 외 arxiv

The proliferation of e-commerce has made web shopping platforms key gateways for customers navigating the vast digital marketplace. Yet this rapid expansion has led to a noisy and fragmented information environment, incr…

DocOps: A Verifiable Benchmark for Autonomous Agents in Complex Document Operations

2026-07-22 · Jiazhen Jiang, Boxi Cao, Lingyong Yan, Yaojie Lu 외 arxiv

As autonomous agents rapidly evolve, their ability to reliably manipulate ubiquitous digital documents has become critical for enabling general-purpose AI assistants and automating complex workspace workflows. In this pa…

FinanceHarness: Autonomous Financial Deep Research Framework

2026-07-30 · Yijia Xiao, Rujun Han, Yanfei Chen, Zifeng Wang 외 arxiv

Powered by advances in LLMs and autonomous agents, deep research has become one of the most widely adopted agentic products. However, most deep research systems write general-purpose reports, which are inadequate for fin…

From Scientific Texts to Verifiable Code: Automating the Process with Transformers

2025-01-09 · Changjie Wang, Mariano Scazzariello, Marco Chiesa

Despite the vast body of research literature proposing algorithms with formal guarantees, the amount of verifiable code in today's systems remains minimal. This discrepancy stems from the inherent difficulty of verifying…

AEMA: Verifiable Evaluation Framework for Trustworthy and Controlled Agentic LLM Systems

2026-01-17 · YenTing Lee, Keerthi Koneru, Zahra Moslemi, Sheethal Kumar 외 arxiv

Evaluating large language model (LLM)-based multi-agent systems remains a critical challenge, as these systems must exhibit reliable coordination, transparent decision-making, and verifiable performance across evolving t…