paper-with-me

홈 › Papers

FlashEvolve: Accelerating Agent Self-Evolution with Asynchronous Stage Orchestration

2026-05-08 · Zhengding Hu, Mingge Lu, Zhen Wang, Jixuan Ruan, Chang Chen, Zaifeng Pan, Yue Guan, Ruiyi Wang, Zhongkai Yu, Chao Zhang, Yufei Ding arxiv

LLM-based evolution has emerged as a promising way to improve agents by refining non-parametric artifacts, but its wall-clock cost remains a major bottleneck. We identify that this cost comes from synchronized stage execution and imbalance inside each LLM-heavy stage. We present FlashEvolve, an efficient framework that replaces synchronized execution with asynchronous workers and queues, allowing different stages and steps to overlap. To handle data staleness introduced by asynchrony, FlashEvolve tracks artifact versions and applies different policies to update, discard, or patch stale artifacts. Unlike weight-space staleness in asynchronous RL, language-space staleness is inspectable and repairable: a stale artifact is not just delayed work, but readable evidence that the LLM can reflect on, revise, and turn into useful evolution signal. FlashEvolve further improves throughput and token efficiency with speculative stage completion and adaptive workflow control. On GEPA workloads, FlashEvolve improves proposal throughput by $3.5\times$ on local vLLM and $4.9\times$ on API serving over synchronous GEPA. The same design also applies to ACE and Meta-Harness.

📄 PDF Abstract BibTeX arXiv:2605.08520

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Massively-concurrent Agent-based Evolutionary Computing

2015-01-27 · D. Krzywicki, W. Turek, A. Byrski, M. Kisiel-Dorohinicki

The fusion of the multi-agent paradigm with evolutionary computation yielded promising results in many optimization problems. Evolutionary multi-agent system (EMAS) are more similar to biological evolution than classical…

Evolutionary Algorithms

Cloud-mediated self-triggered synchronization of a general linear multi-agent system over a directed graph

2023-09-11 · Takumi Namba, Kiyotsugu Takaba

This paper proposes a self-triggered synchronization control method of a general high-order linear time-invariant multi-agent system through a cloud repository. In the cloud-mediated self-triggered control, each agent as…

Part II: ROLL Flash -- Accelerating RLVR and Agentic Training with Asynchrony

2025-10-13 · Han Lu, Zichen Liu, Shaopan Xiong, Yancheng He 외 arxiv

Synchronous Reinforcement Learning (RL) post-training has emerged as a crucial step for enhancing Large Language Models (LLMs) with diverse capabilities. However, many systems designed to accelerate RL post-training stil…

Reinforcement Learning

Darwin Mobile Agent: A Roadmap for Self-Evolution

2026-05-26 · Daniel Beechey, Derek Yuen, Jianheng Liu, Dezhao Luo 외 arxiv

The goal of artificial intelligence is to create agents capable of general, adaptive behaviour in open-ended environments. Guided by the "Bitter Lesson", we argue that the most effective path toward this goal is to syste…

Reinforcement Learning

Robotouille: An Asynchronous Planning Benchmark for LLM Agents

2025-02-06 · Gonzalo Gonzalez-Pumariega, Leong Su Yean, Neha Sunkara, Sanjiban Choudhury

Effective asynchronous planning, or the ability to efficiently reason and plan over states and actions that must happen in parallel or sequentially, is essential for agents that must account for time delays, reason over …

Language ModelingLanguage ModellingLarge Language ModelTask Planning