paper-with-me

Papers

Data augmentation for efficient learning from parametric experts

2022-05-23 · NeurIPS 2021 12 · Alexandre Galashov, Josh Merel, Nicolas Heess

We present a simple, yet powerful data-augmentation technique to enable data-efficient learning from parametric experts for reinforcement and imitation learning. We focus on what we call the policy cloning setting, in which we use online or offline queries of an expert or expert policy to inform the behavior of a student policy. This setting arises naturally in a number of problems, for instance as variants of behavior cloning, or as a component of other algorithms such as DAGGER, policy distillation or KL-regularized RL. Our approach, augmented policy cloning (APC), uses synthetic states to induce feedback-sensitivity in a region around sampled trajectories, thus dramatically reducing the environment interactions required for successful cloning of the expert. We achieve highly data-efficient transfer of behavior from an expert to a student policy for high-degrees-of-freedom control problems. We demonstrate the benefit of our method in the context of several existing and widely used algorithms that include policy cloning as a constituent part. Moreover, we highlight the benefits of our approach in two practically relevant settings (a) expert compression, i.e. transfer to a student with fewer parameters; and (b) transfer from privileged experts, i.e. where the expert has a different observation space than the student, usually including access to privileged information.

📄 PDF Abstract BibTeX arXiv:2205.11448

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationImitation Learning

Similar Papers 제목 키워드 기반

Decoupled Mixture-of-Experts for Parametric Knowledge Injection

2026-06-12 · Baoqing Yue, Weihang Su, Qingyao Ai, Yichen Tang 외 arxiv

Knowledge injection aims to equip large language models (LLMs) with external, domain-specific, or time-sensitive knowledge. Existing approaches typically face a trade-off between flexibility and integration: retrieval-au…

Non Parametric Data Augmentations Improve Deep-Learning based Brain Tumor Segmentation

2021-11-25 · Hadas Ben-Atya, Ori Rajchert, Liran Goshen, Moti Freiman

Automatic brain tumor segmentation from Magnetic Resonance Imaging (MRI) data plays an important role in assessing tumor response to therapy and personalized treatment stratification.Manual segmentation is tedious and su…

Brain Tumor SegmentationData AugmentationDecoderSegmentation+1

Knowledge-in-Context: Towards Knowledgeable Semi-Parametric Language Models

2022-10-28 · Xiaoman Pan, Wenlin Yao, Hongming Zhang, Dian Yu 외

Fully-parametric language models generally require a huge number of model parameters to store the necessary knowledge for solving multiple natural language tasks in zero/few-shot settings. In addition, it is hard to adap…

Common Sense ReasoningCoreference ResolutionLanguage ModelingLanguage Modelling+7

Mask-guided Data Augmentation for Multiparametric MRI Generation with a Rare Hepatocellular Carcinoma

2023-07-30 · Karen Sanchez, Carlos Hinojosa, Kevin Arias, Henry Arguello 외

Data augmentation is classically used to improve the overall performance of deep learning models. It is, however, challenging in the case of medical applications, and in particular for multiparametric datasets. For examp…

AnatomyData AugmentationImage Generation

Zemi: Learning Zero-Shot Semi-Parametric Language Models from Multiple Tasks

2022-10-01 · Zhenhailong Wang, Xiaoman Pan, Dian Yu, Dong Yu 외

Although large language models have achieved impressive zero-shot ability, the huge model size generally incurs high cost. Recently, semi-parametric language models, which augment a smaller language model with an externa…

Language ModelingLanguage ModellingRetrievalText Augmentation+1