paper-with-me

홈 › Papers

Structurally Aligned Subtask-Level Memory for Software Engineering Agents

2026-02-25 · Kangning Shen, Jingyuan Zhang, Chenxi Sun, Wencong Zeng, Yang Yue arxiv

Large Language Models (LLMs) have demonstrated significant potential as autonomous software engineering (SWE) agents. Recent work has further explored augmenting these agents with memory mechanisms to support long-horizon reasoning. However, these approaches typically operate at a coarse instance granularity, treating the entire problem-solving episode as the atomic unit of storage and retrieval. We empirically demonstrate that instance-level memory suffers from a fundamental granularity mismatch, resulting in misguided retrieval when tasks with similar surface descriptions require distinct reasoning logic at specific stages. To address this, we propose Structurally Aligned Subtask-Level Memory, a method that aligns memory storage, retrieval, and updating with the agent's functional decomposition. Extensive experiments on SWE-bench Verified demonstrate that our method consistently outperforms both vanilla agents and strong instance-level memory baselines across diverse backbones, improving mean Pass@1 over the vanilla agent by +4.7 pp on average (e.g., +6.8 pp on Gemini 2.5 Pro). Performance gains grow with more interaction steps, showing that leveraging past experience benefits long-horizon reasoning in complex software engineering tasks.

📄 PDF Abstract BibTeX arXiv:2602.21611

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments

2026-06-11 · Jundong Xu, Qingchuan Li, Jiaying Wu, Yihuai Lan 외 arxiv

Large language model (LLM) agents have achieved strong performance on a wide range of benchmarks, yet most evaluations assume static environments. In contrast, real-world deployment is inherently dynamic, requiring agent…

Bi-HIL: Bilateral Control-Based Multimodal Hierarchical Imitation Learning via Subtask-Level Progress Rate and Keyframe Memory for Long-Horizon Contact-Rich Robotic Manipulation

2026-03-04 · Thanpimon Buamanee, Masato Kobayashi, Yuki Uranishi arxiv

Long-horizon contact-rich robotic manipulation remains challenging due to partial observability and unstable subtask transitions under contact uncertainty. While hierarchical architectures improve temporal reasoning and …

HyperSkill: Self-Evolving LLM Agents via Hypergraph-Structured Skill Memory

2026-08-17 · Ruiyao Xu, Tiankai Yang, Wei-Chieh Huang arxiv

As agentic tasks grow in complexity, LLM agents increasingly rely on experiential memory to reuse procedural knowledge across tasks. Effective memory design must jointly address what to store, how memory is structured an…

UI-Mem: Self-Evolving Experience Memory for Online Reinforcement Learning in Mobile GUI Agents

2026-02-05 · Han Xiao, Guozhi Wang, Hao Wang, Shilong Liu 외 arxiv

Online Reinforcement Learning (RL) offers a promising paradigm for enhancing GUI agents through direct environment interaction. However, its effectiveness is severely hindered by inefficient credit assignment in long-hor…

Reinforcement Learning

Learning Task Decomposition with Ordered Memory Policy Network

2021-03-19 · Yuchen Lu, Yikang Shen, Siyuan Zhou, Aaron Courville 외

Many complex real-world tasks are composed of several levels of sub-tasks. Humans leverage these hierarchical structures to accelerate the learning process and achieve better generalization. In this work, we study the in…

Inductive Bias