paper-with-me

홈 › Papers

El Agente Forjador: Task-Driven Agent Generation for Quantum Simulation

2026-04-16 · Zijian Zhang, Aiwei Yin, Amaan Baweja, Jiaru Bai, Ignacio Gustin, Varinia Bernales, Alán Aspuru-Guzik arxiv

AI for science promises to accelerate the discovery process. The advent of large language models (LLMs) and agentic workflows enables the expediting of a growing range of scientific tasks. However, most of the current generation of agentic systems depend on static, hand-curated toolsets that hinder adaptation to new domains and evolving libraries. We present El Agente Forjador, a multi-agent framework in which universal coding agents autonomously forge, validate, and reuse computational tools through a four-stage workflow of tool analysis, tool generation, task execution, and iterative solution evaluation. Evaluated across 24 tasks spanning quantum chemistry and quantum dynamics on five coding agent setups, we compare three operating modes: zero-shot generation of tools per task, reuse of a curriculum-built toolset, and direct problem-solving with the coding agents as the baseline. We find that our tool generation and reuse framework consistently improves accuracy over the baseline. We also show that reusing a toolset built by a stronger coding agent can reduce API cost and substantially raises the solution quality for weaker coding agents. Case studies further demonstrate that tools forged for different domains can be combined to solve hybrid tasks. Taken together, these results show that LLM-based agents can use their scientific knowledge and coding capabilities to autonomously build reusable scientific tools, pointing toward a paradigm in which agent capabilities are defined by the tasks they are designed to solve rather than by explicitly engineered implementations.

📄 PDF Abstract BibTeX arXiv:2604.14609

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AgentEvolver: Towards Efficient Self-Evolving Agent System

2025-11-13 · Yunpeng Zhai, Shuchang Tao, Cheng Chen, Anni Zou 외 arxiv

Autonomous agents powered by large language models (LLMs) have the potential to significantly enhance human productivity by reasoning, using tools, and executing complex tasks in diverse environments. However, current ap…

Reinforcement Learning

El Agente Estructural: An Artificially Intelligent Molecular Editor

2026-02-04 · Changhyeok Choi, Yunheng Zou, Marcel Müller, Han Hao 외 arxiv

We present El Agente Estructural, a multimodal, natural-language-driven geometry-generation and manipulation agent for autonomous chemistry and molecular modelling. Unlike molecular generation or editing via generative m…

Multimodal Reasoning

Aurora: Unified Video Editing with a Tool-Using Agent

2026-05-18 · Yongsheng Yu, Ziyun Zeng, Zhiyuan Xiao, Zhenghong Zhou 외 arxiv

Recent video editing models have converged on a unified conditioning design: a single diffusion transformer jointly consumes text, source video, and reference images, and one set of weights covers replacement, removal, s…

Style Transfer

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

2026-05-08 · Zhengkang Guo, Yiyang Li, Lin Qiu, Xiaohua Wang 외 arxiv

As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar workflows and short-range interactions. We introduce AgentEscapeBench,…

El Agente: An Autonomous Agent for Quantum Chemistry

2025-05-05 · Yunheng Zou, Austin H. Cheng, Abdulrahman Aldossary, Jiaru Bai 외

Computational chemistry tools are widely used to study the behaviour of chemical phenomena. Yet, the complexity of these tools can make them inaccessible to non-specialists and challenging even for experts. In this work,…

Computational chemistry