paper-with-me

홈 › Papers

ORXE: Orchestrating Experts for Dynamically Configurable Efficiency

2025-05-07 · Qingyuan Wang, Guoxin Wang, Barry Cardiff, Deepu John

This paper presents ORXE, a modular and adaptable framework for achieving real-time configurable efficiency in AI models. By leveraging a collection of pre-trained experts with diverse computational costs and performance levels, ORXE dynamically adjusts inference pathways based on the complexity of input samples. Unlike conventional approaches that require complex metamodel training, ORXE achieves high efficiency and flexibility without complicating the development process. The proposed system utilizes a confidence-based gating mechanism to allocate appropriate computational resources for each input. ORXE also supports adjustments to the preference between inference cost and prediction performance across a wide range during runtime. We implemented a training-free ORXE system for image classification tasks, evaluating its efficiency and accuracy across various devices. The results demonstrate that ORXE achieves superior performance compared to individual experts and other dynamic models in most cases. This approach can be extended to other applications, providing a scalable solution for diverse real-world deployment scenarios.

📄 PDF Abstract BibTeX arXiv:2505.04850

Code (0)

등록된 구현이 없습니다.

Tasks

image-classificationImage Classification

Similar Papers 제목 키워드 기반

Orchestrating Heterogeneous Experts: A Scalable MoE Framework with Anisotropy-Preserving Fusion

2025-11-18 · Ye Liu, Xu Chen, Wuji Chen, Mang Li arxiv

In cross-border e-commerce, search relevance modeling faces the dual challenge of extreme linguistic diversity and fine-grained semantic nuances. Existing approaches typically rely on scaling up a single monolithic Large…

UCB-type Algorithm for Budget-Constrained Expert Learning

2025-10-26 · Ilgam Latypov, Alexandra Suvorikova, Alexey Kroshnin, Alexander Gasnikov 외 arxiv

In many modern applications, a system must dynamically choose between several adaptive learning algorithms that are trained online. Examples include model selection in streaming environments, switching between trading st…

Reinforcement Learning

OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale

2026-02-05 · Jingze Shi, Zhangyang Peng, Yizhang Zhu, Yifan Wu 외 arxiv

Mixture-of-Experts (MoE) architectures are evolving towards finer granularity to improve parameter efficiency. However, existing MoE designs face an inherent trade-off between the granularity of expert specialization and…

Ageing Mitigation and Loss Control Through Ripple Management in Dynamically Reconfigurable Batteries

2022-11-06 · Tomas Kacetl, Jan Kacetl, Nima Tashakor, Stefan M. Goetz

Dynamically reconfigurable batteries merge battery management with output formation in ac and dc batteries, increasing the available charge, power, and life time. However, the combined ripple generated by the load and th…

Management

Dual-Port Dynamically Reconfigurable Battery with Semi-Controlled and Fully-Controlled Outputs

2022-06-03 · N. Tashakor, J. Kacetl, J. Fang, Z. Li 외

Modular multilevel converters (MMC) and cascaded H-bridge (CHB) converters are an established concept in ultra-high voltage systems. In combination with batteries, these circuits allow dynamically changing the series or …