paper-with-me

홈 › Papers

Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces

2026-05-11 · Simon Yu, Derek Chong, Ananjan Nandi, Dilara Soylu, Jiuding Sun, Christopher D Manning, Weiyan Shi arxiv

As LLM agent systems take on more complex tasks, they increasingly rely on meta-agents: higher-order agents that create, operate on and manage other agents. Meta-agent operations such as coordinating agents, halting risky actions before execution, or repairing failed runs, require runtime manipulation of agentic execution. Yet existing agentic substrates make this difficult: they expose only transcripts and environment snapshots, forcing meta-agents to build ad hoc tooling to reconstruct and operate over full execution state. Therefore, we introduce Shepherd, a Python substrate grounded in functional programming principles, where an agent's execution is itself a first-class object that a meta-agent can easily inspect and transform. Every model action, tool call, and environment change becomes a structured event in a reversible, Git-like execution trace, where any past state can be reverted 5x faster than docker commit and fork. Three example use cases show Shepherd's versatility: (1) a supervisor meta-agent prevents conflicts among parallel coding agents, lifting pair-coding pass rate from 28.8% to 54.7% on CooperBench; (2) a counterfactual optimization meta-agent repairs agent workflows by proposing edits and replaying runs from the point of changed behavior, outperforming MetaHarness on Terminal-Bench 2.0 by 12.8% with 58% lower wall-clock; (3) a training meta-agent picks fork points during rollouts to improve credit assignment in long-horizon agentic RL, doubling GRPO's uplift on Terminal-Bench 2.0. We open-source Shepherd to enable principled and efficient operations over agentic execution for both users and meta-agents.

📄 PDF Abstract BibTeX arXiv:2605.10913

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GUI-Shepherd: Reliable Process Reward and Verification for Long-Sequence GUI Tasks

2025-09-28 · Cong Chen, Kaixiang Ji, Hao Zhong, Muzhi Zhu 외 arxiv

Autonomous agents for long-sequence Graphical User Interface tasks are hindered by sparse rewards and the intractable credit assignment problem. To address these challenges, we introduce GUI-Shepherd, a Process Reward Mo…

A Comprehensive Review of Shepherding as a Bio-inspired Swarm-Robotics Guidance Approach

2019-12-17 · Nathan K Long, Karl Sammut, Daniel Sgarioto, Matthew Garratt 외

The simultaneous control of multiple coordinated robotic agents represents an elaborate problem. If solved, however, the interaction between the agents can lead to solutions to sophisticated problems. The concept of swar…

Communication-Free Shepherding Navigation with Multiple Steering Agents

2022-05-17 · Aiyi Li, Masaki Ogura, Naoki Wakamiya

Swarm guidance addresses a challenging problem considering the navigation and control of a group of passive agents. To solve this problem, shepherding offers a bio-inspired technique of navigating such group of agents by…

Web-Shepherd: Advancing PRMs for Reinforcing Web Agents

2025-05-21 · Hyungjoo Chae, Sunghwan Kim, Junhee Cho, Seungone Kim 외

Web navigation is a unique domain that can automate many repetitive real-life tasks and is challenging as it requires long-horizon sequential decision making beyond typical multimodal large language model (MLLM) tasks. Y…

Large Language ModelMultimodal Large Language ModelSequential Decision Making

Transparent Machine Education of Neural Networks for Swarm Shepherding Using Curriculum Design

2019-01-04 · Alexander Gee, Hussein Abbass

Swarm control is a difficult problem due to the need to guide a large number of agents simultaneously. We cast the problem as a shepherding problem, similar to biological dogs guiding a group of sheep towards a goal. The…