paper-with-me

Papers

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

2026-05-28 · Jiajie Fu, Junwen Chen, Mengzhao Wang, Aoxiang He, Maojia Sheng, Xiangyu Ke, Yifan Zhu, Yunjun Gao arxiv

Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for maintaining stateful, long-term interactions. Existing memory approaches often rely on simplistic extraction methods that lead to incomplete memories or use rigid, single-purpose memory extraction prompts tailored to a single use case, such as chatbots. Consequently, they lack generalizability and perform poorly across diverse downstream tasks. To bridge this gap, we introduce the Memory Base, a novel data management paradigm for managing the persistent state of long-term interactions. It is characterized by three core principles: selective extraction of high-value memories from raw information streams; inherent statefulness and evolution, where memory content is progressively summarized, corrected, and temporally weighted to prioritize recent interactions; and a generalizable abstraction paradigm designed for robust transferability across diverse applications, including education, recommendation, and agent memory. Building on this foundation, we present VikingMem, an end-to-end Memory Base Management System implemented on the VikingDB vector engine. VikingMem materializes this paradigm through interconnected event and entity abstractions. It features event-centric memory extraction to selectively handle complex information streams, while entities are dynamically updated by events to achieve stateful evolution. Using temporal compression via a topic-wise timeline and time-weighted recall, the system progressively produces high-level summary memories, prioritizes recent items, and compresses and fades older ones. Extensive evaluations on long-term memory benchmarks demonstrate that VikingMem outperformes baselines by up to 30% in memory retrieval effectiveness while maintaining the low latency essential for interactive applications.

📄 PDF Abstract BibTeX arXiv:2605.29640

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments

2026-06-04 · Parth Asawa, Christopher M. Glaze, Gabriel Orlanski, Ramya Ramakrishnan 외 arxiv

Continual learning, the ability of AI systems to improve through sequential experience, has attracted substantial interest, but no high-quality benchmark exists to evaluate it. We introduce Continual Learning Bench (CL-B…

Continual Learning

Stateful KV Cache Management for LLMs: Balancing Space, Time, Accuracy, and Positional Fidelity

2025-10-23 · Pratik Poudel arxiv

The Key-Value (KV) cache is integral to efficient autoregressive inference in large language models (LLMs), yet its unbounded growth in stateful multi-turn scenarios presents major challenges. This paper examines the int…

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads

2026-06-04 · Yasmine Omri, Ziyu Gan, Zachary Broveak, Robin Geens 외 arxiv

LLM agents are increasingly deployed on long-horizon tasks requiring sustained reasoning over extended interaction histories. Realizing this at scale requires agents to persistently store, retrieve, and update their own …

MemForest: An Efficient Agent Memory System with Hierarchical Temporal Indexing

2026-05-16 · Han Chen, Zining Zhang, Wenqi Pei, Bingsheng He 외 arxiv

Memory is a fundamental component for enabling long-context LLM agents, supporting persistent state across interactions through a continuous serve-and-update lifecycle. Despite substantial prior work, existing systems su…

StateLinFormer: Stateful Training Enhancing Long-term Memory in Navigation

2026-03-24 · Zhiyuan Chen, Yuxuan Zhong, Fan Wang, Bo Yu 외 arxiv

Effective navigation intelligence relies on long-term memory to support both immediate generalization and sustained adaptation. However, existing approaches face a dilemma: modular systems rely on explicit mapping but la…