paper-with-me

홈 › Papers

Mixture of Experts in a Mixture of RL settings

2024-06-26 · Timon Willi, Johan Obando-Ceron, Jakob Foerster, Karolina Dziugaite, Pablo Samuel Castro

Mixtures of Experts (MoEs) have gained prominence in (self-)supervised learning due to their enhanced inference efficiency, adaptability to distributed training, and modularity. Previous research has illustrated that MoEs can significantly boost Deep Reinforcement Learning (DRL) performance by expanding the network's parameter count while reducing dormant neurons, thereby enhancing the model's learning capacity and ability to deal with non-stationarity. In this work, we shed more light on MoEs' ability to deal with non-stationarity and investigate MoEs in DRL settings with "amplified" non-stationarity via multi-task training, providing further evidence that MoEs improve learning capacity. In contrast to previous work, our multi-task results allow us to better understand the underlying causes for the beneficial effect of MoE in DRL training, the impact of the various MoE components, and insights into how best to incorporate them in actor-critic-based DRL networks. Finally, we also confirm results from previous work.

📄 PDF Abstract BibTeX arXiv:2406.18420

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement LearningMixture-of-ExpertsSelf-Supervised Learning

Methods 이 논문이 사용한 방법론

MoE 설명 없음

Similar Papers 제목 키워드 기반

A similarity-based Bayesian mixture-of-experts model

2020-12-03 · Tianfang Zhang, Rasmus Bokrantz, Jimmy Olsson

We present a new nonparametric mixture-of-experts model for multivariate regression problems, inspired by the probabilistic k-nearest neighbors algorithm. Using a conditionally specified model, predictions for out-of-sam…

Mixture-of-Expertsmodel

Mixture of LoRA Experts for Low-Resourced Multi-Accent Automatic Speech Recognition

2025-05-26 · Raphaël Bagat, Irina Illina, Emmanuel Vincent

We aim to improve the robustness of Automatic Speech Recognition (ASR) systems against non-native speech, particularly in low-resourced multi-accent settings. We introduce Mixture of Accent-Specific LoRAs (MAS-LoRA), a f…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Mixture of Length and Pruning Experts for Knowledge Graphs Reasoning

2025-07-28 · Enjun Du, Siyi Liu, Yongqi Zhang arxiv

Knowledge Graph (KG) reasoning, which aims to infer new facts from structured knowledge repositories, plays a vital role in Natural Language Processing (NLP) systems. Its effectiveness critically depends on constructing …

Knowledge Graphs

Peirce in the Machine: How Mixture of Experts Models Perform Hypothesis Construction

2024-06-24 · Bruce Rushing

Mixture of experts is a prediction aggregation method in machine learning that aggregates the predictions of specialized experts. This method often outperforms Bayesian methods despite the Bayesian having stronger induct…

Mixture-of-Experts

Mixtures of Gaussian Process Experts with SMC$^2$

2022-08-26 · Teemu Härkönen, Sara Wade, Kody Law, Lassi Roininen

Gaussian processes are a key component of many flexible statistical and machine learning models. However, they exhibit cubic computational complexity and high memory constraints due to the need of inverting and storing a…

Gaussian Processes