paper-with-me

홈 › Papers

When Model Merging Rivals Joint Multi-Task Reinforcement Learning: A Task-Vector Geometry Analysis

2026-07-17 · S. Aaron McClendon arxiv

Model merging is promoted as a substitute for joint multi-task training, yet in the reinforcement-learning setting this substitution is essentially never tested against the baseline it claims to replace: methods merge independently released agents precisely because a joint model is unavailable. We build the missing comparison. Training difficulty-1 and difficulty-2 Qwen3-8B specialists on the AppWorld agent benchmark with LOOP, we merge them (TIES, RAM+) and pit the result against a jointly trained model on the same data. On task-goal completion, merging matches joint RL -- and every merge variant is statistically indistinguishable. To explain why merge method does not matter here, we measure the geometry of the specialists' task vectors, which carries no task-sampling noise: they are near-orthogonal (cosine 0.06 - 0.10) despite ~65% support overlap, a small, shared direction that grows over training and that we calibrate against a random-init floor and a same-run ceiling to confirm it reflects learning, not the low-rank parameterization. Because direction and support are decoupled, support and sign-based merging (RAM, TIES) collapse to near-uniform averaging. We release all code and statistics.

📄 PDF Abstract BibTeX arXiv:2607.16062

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Single-Input Multi-Output Model Merging: Leveraging Foundation Models for Dense Multi-Task Learning

2025-04-15 · Juan Garcia Giraldo, Nikolaos Dimitriadis, Ke Wang, Pascal Frossard

Model merging is a flexible and computationally tractable approach to merge single-task checkpoints into a multi-task model. Prior work has solely focused on constrained multi-task settings where there is a one-to-one ma…

Multi-Task LearningScene UnderstandingTask Arithmetic

Joint Batching and Scheduling for High-Throughput Multiuser Edge AI with Asynchronous Task Arrivals

2023-07-15 · Yihan Cang, Ming Chen, Kaibin Huang

In this paper, we study joint batching and (task) scheduling to maximise the throughput (i.e., the number of completed tasks) under the practical assumptions of heterogeneous task arrivals and deadlines. The design aims …

BenchmarkingScheduling

Channel Estimation for Ambient Backscatter Communication Systems with Massive-Antenna Reader

2019-06-30

Ambient backscatter, an emerging green communication technology, has aroused great interest from both academia and industry. One open problem for ambient backscatter communication (AmBC) systems is channel estimation for…

Multi-Turn Reasoning LLMs for Task Offloading in Mobile Edge Computing

2026-04-08 · Ning Yang, Chuangxin Cheng, Haijun Zhang arxiv

Emerging computation-intensive applications impose stringent latency requirements on resource-constrained mobile devices. Mobile Edge Computing (MEC) addresses this challenge through task offloading. However, designing e…

Reinforcement Learning

A Hybrid Game-Theory and Deep Learning Framework for Predicting Tourist Arrivals via Big Data Analytics and Opinion Leader Detection

2025-07-04 · Ali Nikseresht arxiv

In the era of Industry 5.0, data-driven decision-making has become indispensable for optimizing systems across Industrial Engineering. This paper addresses the value of big data analytics by proposing a novel non-linear …