paper-with-me

Papers

SwitchMT: An Adaptive Context Switching Methodology for Scalable Multi-Task Learning in Intelligent Autonomous Agents

2025-04-18 · Avaneesh Devkota, Rachmad Vidya Wicaksana Putra, Muhammad Shafique

The ability to train intelligent autonomous agents (such as mobile robots) on multiple tasks is crucial for adapting to dynamic real-world environments. However, state-of-the-art reinforcement learning (RL) methods only excel in single-task settings, and still struggle to generalize across multiple tasks due to task interference. Moreover, real-world environments also demand the agents to have data stream processing capabilities. Toward this, a state-of-the-art work employs Spiking Neural Networks (SNNs) to improve multi-task learning by exploiting temporal information in data stream, while enabling lowpower/energy event-based operations. However, it relies on fixed context/task-switching intervals during its training, hence limiting the scalability and effectiveness of multi-task learning. To address these limitations, we propose SwitchMT, a novel adaptive task-switching methodology for RL-based multi-task learning in autonomous agents. Specifically, SwitchMT employs the following key ideas: (1) a Deep Spiking Q-Network with active dendrites and dueling structure, that utilizes task-specific context signals to create specialized sub-networks; and (2) an adaptive task-switching policy that leverages both rewards and internal dynamics of the network parameters. Experimental results demonstrate that SwitchMT achieves superior performance in multi-task learning compared to state-of-the-art methods. It achieves competitive scores in multiple Atari games (i.e., Pong: -8.8, Breakout: 5.6, and Enduro: 355.2) compared to the state-of-the-art, showing its better generalized learning capability. These results highlight the effectiveness of our SwitchMT methodology in addressing task interference while enabling multi-task learning automation through adaptive task switching, thereby paving the way for more efficient generalist agents with scalable multi-task learning capabilities.

📄 PDF Abstract BibTeX arXiv:2504.13541

Code (0)

등록된 구현이 없습니다.

Tasks

Atari GamesMulti-Task LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Sustainable LLM Inference using Context-Aware Model Switching

2026-02-25 · Yuvarani, Akashdeep Singh, Zahra Fathanah, Salsabila Harlen 외 arxiv

Large language models have become central to many AI applications, but their growing energy consumption raises serious sustainability concerns. A key limitation in current AI deployments is the reliance on a one-size-fit…

Optimal Transmission Switching and Busbar Splitting in Hybrid AC/DC Grids

2024-11-29 · Giacomo Bastianel, Marta Vanin, Dirk Van Hertem, Hakan Ergun

Driven by global climate goals, an increasing amount of Renewable Energy Sources (RES) is currently being installed worldwide. Especially in the context of offshore wind integration, hybrid AC/DC grids are considered to …

On the Design of Limit Cycles of Planar Switching Affine Systems

2023-03-29 · Nils Hanke, Olaf Stursberg

In the context of studying periodic processes, this paper investigates first under which conditions switching affine systems in the plane generate stable limit cycles. Based on these conditions, a design methodology is p…

Real-time Hybrid System Identification with Online Deterministic Annealing

2024-08-03 · Christos Mavridis, Karl Henrik Johansson

We introduce a real-time identification method for discrete-time state-dependent switching systems in both the input--output and state-space domains. In particular, we design a system of adaptive algorithms running in tw…

Computational Efficiency

PATS: Process-Level Adaptive Thinking Mode Switching

2025-05-25 · Yi Wang, Junxiao Liu, Shimao Zhang, Jiajun Chen 외

Current large-language models (LLMs) typically adopt a fixed reasoning strategy, either simple or complex, for all questions, regardless of their difficulty. This neglect of variation in task and reasoning process comple…

Computational Efficiency