paper-with-me

홈 › Papers

TSN-Affinity: Similarity-Driven Parameter Reuse for Continual Offline Reinforcement Learning

2026-04-28 · Dominik Żurek, Kamil Faber, Marcin Pietron, Paweł Gajewski, Roberto Corizzo arxiv

Continual offline reinforcement learning (CORL) aims to learn a sequence of tasks from datasets collected over time while preserving performance on previously learned tasks. This setting corresponds to domains where new tasks arise over time, but adapting the model in live environment interactions is expensive, risky, or impossible. However, CORL inherits the dual difficulty of offline reinforcement learning and adapting while preventing catastrophic forgetting. Replay-based continual learning approaches remain a strong baseline but incur memory overhead and suffer from a distribution mismatch between replayed samples and newly learned policies. At the same time, architectural continual learning methods have shown strong potential in supervised learning but remain underexplored in CORL. In this work, we propose TSN-Affinity, a novel CORL method based on TinySubNetworks and Decision Transformer. The method enables task-specific parameterization and controlled knowledge sharing through a RL-aware reuse strategy that routes tasks according to action compatibility and latent similarity. We evaluate the approach on benchmarks based on Atari games and simulations of manipulation tasks with the Franka Emika Panda robotic arm, covering both discrete and continuous control. Results show strong retention from sparse SubNetworks, with routing further improving multi-task performance. Our findings suggest that similarity-guided architectural reuse is a strong and viable alternative to replay-based strategies in a CORL setting. Our code is available at: https://github.com/anonymized-for-submission123/tsn-affinity.

📄 PDF Abstract BibTeX arXiv:2604.25898

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningContinuous ControlContinual LearningAtari Games

Similar Papers 제목 키워드 기반

Continual Sequence Generation with Adaptive Compositional Modules

2022-03-20 · ACL 2022 5 · Yanzhe Zhang, Xuezhi Wang, Diyi Yang

Continual learning is essential for real-world deployment when there is a need to quickly adapt the model to new tasks without forgetting knowledge of old tasks. Existing work on continual sequence generation either alwa…

Continual LearningTransfer Learning

Similarity-based context aware continual learning for spiking neural networks

2024-10-28 · Bing Han, Feifei Zhao, Yang Li, Qingqun Kong 외

Biological brains have the capability to adaptively coordinate relevant neuronal populations based on the task context to learn continuously changing tasks in real-world environments. However, existing spiking neural net…

class-incremental learningClass Incremental LearningContinual LearningIncremental Learning

iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models

2026-05-19 · Xuezhi Cui, Dongbo Zhou, Wang Guo, Zeyuan Wang 외 arxiv

Vision-Language Models require efficient adaptation to continually emerging downstream tasks. While Parameter-Efficient Fine-Tuning mitigates catastrophic forgetting, assigning isolated modules per task leads to paramete…

parameter-efficient fine-tuningContinual Learning

Toward Sustainable Continual Learning: Detection and Knowledge Repurposing of Similar Tasks

2022-10-11 · Sijia Wang, Yoojin Choi, Junya Chen, Mostafa El-Khamy 외

Most existing works on continual learning (CL) focus on overcoming the catastrophic forgetting (CF) problem, with dynamic models and replay methods performing exceptionally well. However, since current works tend to assu…

Continual Learning

Conformal Prediction based Spectral Clustering

2019-09-17 · Lalith Srikanth Chintalapati, Raghunatha Sarma Rachakonda

Spectral Clustering(SC) is a prominent data clustering technique of recent times which has attracted much attention from researchers. It is a highly data-driven method and makes no strict assumptions on the structure of …

ClusteringConformal PredictionPrediction