paper-with-me

홈 › Papers

Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation

2026-02-02 · Zhanghao Hu, Qinglin Zhu, Runcong Zhao, Di Liang, Hanqi Yan, Yulan He, Lin Gui arxiv

Standard Retrieval Augmented Generation (RAG) is poorly matched to agent memory. Unlike large heterogeneous corpora, agent memory forms a bounded and coherent interaction stream in which many spans are highly correlated or near duplicates. As a result, flat top-$k$ similarity retrieval often returns redundant context, while summary-centric hierarchies can blur the subtle details that distinguish one candidate from another. We argue that agent memory should follow the principle of decoupling before aggregation: the system should first isolate reusable facts, updates, and distinguishing details from similar histories, and only then organise them for efficient retrieval. Based on this principle, we propose xMemory, which constructs a revisable hierarchical memory structure from original messages to segments, memory components, and groups. xMemory segments interaction history into local events, decouples each segment into memory components, aggregates related components into high-level groups using a sparsity--semantic faithfulness objective, and maintains this structure incrementally as memory evolves. At inference time, xMemory retrieves top-down, first selecting a compact backbone of complementary groups and components, and then expanding to segments and raw messages only when additional evidence reduces the reader's uncertainty. Experiments on LoCoMo and PerLTQA across diverse open source and closed source LLMs show consistent gains in answer quality and inference token efficiency, supported by analyses of redundancy, evidence density, and coverage.

📄 PDF Abstract BibTeX arXiv:2602.02007

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MAGMA: A Multi-Graph based Agentic Memory Architecture for AI Agents

2026-01-06 · Dongming Jiang, Yi Li, Guanpeng Li, Bingzhe Li arxiv

Memory-Augmented Generation (MAG) extends Large Language Models with external memory to support long-context reasoning, but existing approaches largely rely on semantic similarity over monolithic memory stores, entanglin…

Semantic Similarity

Beyond Retrieval: Analytic Memory for Multimodal Agents

2026-07-31 · Zhoujin Tian, Hao Zhang, Yao Tian, Cheng Chen 외 arxiv

Long-term multimodal memory must support not only retrieving relevant information but also computing over observations accumulated across interactions. Existing systems largely emphasize \emph{retrieval memory}, organizi…

MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents

2026-05-27 · Zihan Li, Xingyu Fan, Feifei Li, Wenhui Que arxiv

Existing agent memory systems universally follow what we term a Memory-as-Tool paradigm where a single query triggers one-shot retrieval of flat passage lists, suffering from passive invocation, reasoning-retrieval decou…

CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning

2026-01-21 · Zhiyuan Lu, Chenliang Li, Yingcheng Shi, Weizhou Shen 외 arxiv

While large language models now handle million-token contexts, their capacity for reasoning across entire document repositories remains largely untested. Existing benchmarks are inadequate, as they are mostly limited to …

UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization

2026-04-15 · Zhengxi Lu, Fei Tang, Guangyi Liu, Kaitao Song 외 arxiv

MLLM-based GUI agents have demonstrated strong capabilities in complex user interface interaction tasks. However, long-horizon scenarios remain challenging, as these agents are burdened with tasks beyond their intrinsic …