paper-with-me

홈 › Papers

IR3DE: A Linear Router for Large Language Models

2026-06-04 · Eros Fanì, Oğuzhan Ersoy arxiv

Foundational Large Language Models (LLMs) demonstrate proficiency on a wide range of general tasks, and achieve remarkable results on various specialized tasks via domain-expert LLMs. With the ever-growing list of available LLMs, inference routers are being proposed to select the most appropriate LLM for each prompt. However, existing routing methods either optimize cost across weak-to-strong generalist LLMs or require substantial training to support domain-expertise routing. In this paper, we propose IR3DE, a Ridge Regression-based Router for Domain Experts that provides cheap and fast routing decisions for each prompt. We evaluate IR3DE in two Causal Language Modeling (CLM) settings where the tasks are next-token prediction for all domains, and one reasoning setting where each domain has its own distinct reasoning task. Despite being a linear router, IR3DE achieves performance comparable to the other baselines in both CLM settings, and surpassing them in the reasoning setting, with a normalized performance of 98.4%. Moreover, IR3DE enables the addition or removal of new domain experts without requiring the router to be retrained from scratch, allowing a dynamic set of LLMs to be served with minimal disruption to the router itself. Our code is available at: github.com/gensyn-ai/IR3DE.

📄 PDF Abstract BibTeX arXiv:2606.06098

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RouterDC: Query-Based Router by Dual Contrastive Learning for Assembling Large Language Models

2024-09-30 · Shuhao Chen, Weisen Jiang, Baijiong Lin, James T. Kwok 외

Recent works show that assembling multiple off-the-shelf large language models (LLMs) can harness their complementary abilities. To achieve this, routing is a promising method, which learns a router to select the most su…

Contrastive Learning

EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models

2025-05-28 · Linglin Jing, Yuting Gao, Zhigang Wang, Wang Lan 외

Recent advancements have shown that the Mixture of Experts (MoE) approach significantly enhances the capacity of large language models (LLMs) and improves performance on downstream tasks. Building on these promising resu…

Mixture-of-ExpertsMMETextVQA

Statistical Advantages of Perturbing Cosine Router in Mixture of Experts

2024-05-23 · Huy Nguyen, Pedram Akbarian, Trang Pham, Trang Nguyen 외

The cosine router in Mixture of Experts (MoE) has recently emerged as an attractive alternative to the conventional linear router. Indeed, the cosine router demonstrates favorable performance in image and language tasks …

Mixture-of-Experts

Life-Cycle Routing Vulnerabilities of LLM Router

2025-03-09 · Qiqi Lin, Xiaoyang Ji, Shengfang Zhai, Qingni Shen 외

Large language models (LLMs) have achieved remarkable success in natural language processing, yet their performance and computational costs vary significantly. LLM routers play a crucial role in dynamically balancing the…

Adversarial Robustness

Exploring Domain Robust Lightweight Reward Models based on Router Mechanism

2024-07-24 · Hyuk Namgoong, Jeesu Jung, SangKeun Jung, YoonHyung Roh

Recent advancements in large language models have heavily relied on the large reward model from reinforcement learning from human feedback for fine-tuning. However, the use of a single reward model across various domains…

Language ModelingLanguage ModellingMixture-of-ExpertsSmall Language Model