paper-with-me

홈 › Papers

PaCo: Parameter-Compositional Multi-Task Reinforcement Learning

2022-10-21 · Lingfeng Sun, Haichao Zhang, Wei Xu, Masayoshi Tomizuka

The purpose of multi-task reinforcement learning (MTRL) is to train a single policy that can be applied to a set of different tasks. Sharing parameters allows us to take advantage of the similarities among tasks. However, the gaps between contents and difficulties of different tasks bring us challenges on both which tasks should share the parameters and what parameters should be shared, as well as the optimization challenges due to parameter sharing. In this work, we introduce a parameter-compositional approach (PaCo) as an attempt to address these challenges. In this framework, a policy subspace represented by a set of parameters is learned. Policies for all the single tasks lie in this subspace and can be composed by interpolating with the learned set. It allows not only flexible parameter sharing but also a natural way to improve training. We demonstrate the state-of-the-art performance on Meta-World benchmarks, verifying the effectiveness of the proposed approach.

📄 PDF Abstract BibTeX arXiv:2210.11653

Code (1)

ttotmoon/paco-mtrl 공식 구현 pytorch

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

DeepACO: Neural-enhanced Ant Systems for Combinatorial Optimization

2023-09-25 · NeurIPS 2023 11 · Haoran Ye, Jiarui Wang, Zhiguang Cao, Helan Liang 외

Ant Colony Optimization (ACO) is a meta-heuristic algorithm that has been successfully applied to various Combinatorial Optimization Problems (COPs). Traditionally, customizing ACO for a specific problem requires the exp…

Combinatorial OptimizationDeep Reinforcement Learning

PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling

2025-12-02 · Bowen Ping, Chengyou Jia, Minnan Luo, Changliang Xia 외 arxiv

Consistent image generation requires faithfully preserving identities, styles, and logical coherence across multiple images, which is essential for applications such as storytelling and character design. Supervised train…

Reinforcement LearningImage Generation

Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning

2023-11-13 · Arjun Bhardwaj, Jonas Rothfuss, Bhavya Sukhija, Yarden As 외

We introduce PACOH-RL, a novel model-based Meta-Reinforcement Learning (Meta-RL) algorithm designed to efficiently adapt control policies to changing dynamics. PACOH-RL meta-learns priors for the dynamics model, allowing…

Meta-LearningMeta Reinforcement Learningreinforcement-learningUncertainty Quantification

PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning

2026-01-09 · Jingcheng Hu, Yinmin Zhang, Shijie Shang, Xiaobo Yang 외 arxiv

We introduce Parallel Coordinated Reasoning (PaCoRe), a training-and-inference framework designed to overcome a central limitation of contemporary language models: their inability to scale test-time compute (TTC) far bey…

Reinforcement Learning

Generalized Parametric Contrastive Learning

2022-09-26 · Jiequan Cui, Zhisheng Zhong, Zhuotao Tian, Shu Liu 외

In this paper, we propose the Generalized Parametric Contrastive Learning (GPaCo/PaCo) which works well on both imbalanced and balanced data. Based on theoretical analysis, we observe that supervised contrastive loss ten…

Contrastive LearningDomain GeneralizationImage ClassificationLong-tail Learning+1