paper-with-me

홈 › Papers

Active Long Term Memory Networks

2016-06-07 · Tommaso Furlanello, Jiaping Zhao, Andrew M. Saxe, Laurent Itti, Bosco S. Tjan

Continual Learning in artificial neural networks suffers from interference and forgetting when different tasks are learned sequentially. This paper introduces the Active Long Term Memory Networks (A-LTM), a model of sequential multi-task deep learning that is able to maintain previously learned association between sensory input and behavioral output while acquiring knew knowledge. A-LTM exploits the non-convex nature of deep neural networks and actively maintains knowledge of previously learned, inactive tasks using a distillation loss. Distortions of the learned input-output map are penalized but hidden layers are free to transverse towards new local optima that are more favorable for the multi-task objective. We re-frame the McClelland's seminal Hippocampal theory with respect to Catastrophic Inference (CI) behavior exhibited by modern deep architectures trained with back-propagation and inhomogeneous sampling of latent factors across epochs. We present empirical results of non-trivial CI during continual learning in Deep Linear Networks trained on the same task, in Convolutional Neural Networks when the task shifts from predicting semantic to graphical factors and during domain adaptation from simple to complex environments. We present results of the A-LTM model's ability to maintain viewpoint recognition learned in the highly controlled iLab-20M dataset with 10 object categories and 88 camera viewpoints, while adapting to the unstructured domain of Imagenet with 1,000 object categories.

📄 PDF Abstract BibTeX arXiv:1606.02355

Code (0)

등록된 구현이 없습니다.

Tasks

Continual LearningDomain Adaptation

Similar Papers 제목 키워드 기반

Bridging the Long-Term Gap: A Memory-Active Policy for Multi-Session Task-Oriented Dialogue

2025-05-26 · Yiming Du, Bingbing Wang, Yang He, Bin Liang 외

Existing Task-Oriented Dialogue (TOD) systems primarily focus on single-session dialogues, limiting their effectiveness in long-term memory augmentation. To address this challenge, we introduce a MS-TOD dataset, the firs…

MemGround: Long-Term Memory Evaluation Kit for Large Language Models in Gamified Scenarios

2026-03-23 · Yihang Ding, Wanke Xia, Yiting Zhao, Jinbo Su 외 arxiv

Current evaluations of long-term memory in LLMs are fundamentally static. By fixating on simple retrieval and short-context inference, they neglect the multifaceted nature of complex memory systems, such as dynamic state…

Explore with Long-term Memory: A Benchmark and Multimodal LLM-based Reinforcement Learning Framework for Embodied Exploration

2026-01-11 · Sen Wang, Bangwei Liu, Zhenkun Gao, Lizhuang Ma 외 arxiv

An ideal embodied agent should possess lifelong learning capabilities to handle long-horizon and complex tasks, enabling continuous operation in general environments. This not only requires the agent to accurately accomp…

Reinforcement LearningQuestion Answering

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory

2026-04-09 · Zhifei Xie, Zongzheng Hu, Fangda Ye, Xin Zhang 외 arxiv

Proactivity is a core expectation for AGI. Prior work remains largely confined to laboratory settings, leaving a clear gap in real-world proactive agent: depth, complexity, ambiguity, precision and real-time constraints.…

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

2026-04-09 · Jingyi Kang, Chunyu Li, Ding Chen, Bo Tang 외 arxiv

Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction as a one-shot, passive transcription from context to structured ent…