paper-with-me

홈 › Papers

Online Reinforcement Learning with Passive Memory

2024-10-18 · Anay Pattanaik, Lav R. Varshney

This paper considers an online reinforcement learning algorithm that leverages pre-collected data (passive memory) from the environment for online interaction. We show that using passive memory improves performance and further provide theoretical guarantees for regret that turns out to be near-minimax optimal. Results show that the quality of passive memory determines sub-optimality of the incurred regret. The proposed approach and results hold in both continuous and discrete state-action spaces.

📄 PDF Abstract BibTeX arXiv:2410.14665

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Interactive Memory Learning for Long-Term Conversations

2026-09-15 · Cai Ke, Jiangyue Yan, Han Zhang, Xin Liu 외 arxiv

Recent advancements in large language models have significantly enhanced the capabilities of agents in modeling long-term conversations. Despite these successes, existing approaches typically adopt a static heuristic par…

Reinforcement LearningTest-time Adaptation

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States

2026-06-30 · Chen Liu, Ling Chen, Hanzhang Zhou, Xu Zhang 외 arxiv

Mobile GUI agents increasingly face long-horizon tasks that require reading, updating, and reusing task-relevant data across pages and applications. Existing methods treat memory largely as passive storage, where past ob…

Reinforcement Learning

EventMemAgent: Hierarchical Event-Centric Memory for Online Video Understanding with Adaptive Tool Use

2026-02-17 · Siwei Wen, Zhangcheng Wang, Xingjian Zhang, Lei Huang 외 arxiv

Online video understanding requires models to perform continuous perception and long-range reasoning within potentially infinite visual streams. Its fundamental challenge lies in the conflict between the unbounded nature…

Reinforcement Learning

From Passive Retrieval to Active Memory Navigation: Learning to Use Memory as a Structured Action Space

2026-07-07 · Yue Xu, Yutao Sun, Yihao Liu, Mengyu Zhou 외 arxiv

Long-term user memory is essential for personalized conversational agents, yet many memory systems still expose memory through passive retrieval interfaces, making the model a consumer of pre-selected evidence. We introd…

Reinforcement Learning

Memory Lens: How Much Memory Does an Agent Use?

2016-11-21 · Christoph Dann, Katja Hofmann, Sebastian Nowozin

We propose a new method to study the internal memory used by reinforcement learning policies. We estimate the amount of relevant past information by estimating mutual information between behavior histories and the curren…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)