paper-with-me

홈 › Papers

Easy Bayesian Transfer Learning with Informative Priors

2023-09-21 · NeurIPS 2023 11

REPRODUCIBILITY SUMMARY Scope of Reproducibility In this work, we study the reproducibility of the paper: Pre-Train Your Loss: Easy Bayesian Transfer Learning with Informative Priors. The paper proposes a three-step pipeline for replacing standard transfer learning with a pre-trained prior. The first step is training a prior, the second is re-scaling of a prior, and the third is inference. The authors claim that increasing the rank and the scaling factor improves performance on the downstream task. They also argue that using Bayesian learning with informative prior leads to a more data-efficient and improved performance compared to standard SGD transfer learning or using non-informative prior. We reproduce the main claims on one of the four data sets in the paper. Methodology We used a combination of the authors' and our code. The authors provided a training pipeline for the user but not the code to fully reproduce the paper. We modified the training pipeline to suit our needs and created a testing pipeline to evaluate the models. We reproduced the results for the Oxford-102-Flowers data set on an Nvidia RTX 3070 GPU using approximately 310 GPU hours for the main results. Results Our results confirm most of the claims tested, although we could not achieve the exact same accuracy due to missing hyper-parameters. We reproduced the trend in how scaling the prior impacts the performance and how a learned prior outperforms a non-learned prior. On contrary, we could not reproduce the effect of rank in low-rank covariance approximation on model performance, as well as the beneficial boost in performance of Bayesian learning compared to the standard SGD. What was easy The authors' implementation provides various training and logging parameters. It is also helpful that the authors provided both the learned priors and scripts for the download, split and pre-processing of the data sets used in the study. What was difficult Setting the environment for the used packages to work correctly was difficult. Although many parameters are available for running the pipeline, their descriptions are misguiding, therefore a lot of time went into clarifying the parameter function and debugging different settings. The training also took a while, especially when training 5 models per data point. Communication with original authors We contacted the authors via e-mail about their pipeline and their use of hyper-parameters but did not hear back.Paper Url: https://openreview.net/forum?id=ao30zaT3YLPaper Review Url: https://openreview.net/forum?id=ao30zaT3YLPaper Venue: ICML 2022

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SGD Stochastic Gradient Descent is an iterative optimization technique that uses minibatches of data to form an expectation of the gradient, rather than the full gradient using…

Similar Papers 제목 키워드 기반

Pre-Train Your Loss: Easy Bayesian Transfer Learning with Informative Priors

2022-05-20 · Ravid Shwartz-Ziv, Micah Goldblum, Hossein Souri, Sanyam Kapoor 외

Deep learning is increasingly moving towards a transfer learning paradigm whereby large foundation models are fine-tuned on downstream tasks, starting from an initialization learned on the source task. But an initializat…

Deep LearningTransfer Learning

PAC-Bayesian Policy Evaluation for Reinforcement Learning

2012-02-14 · Mahdi Milani Fard, Joelle Pineau, Csaba Szepesvari

Bayesian priors offer a compact yet general means of incorporating domain knowledge into many learning tasks. The correctness of the Bayesian analysis and inference, however, largely depends on accuracy and correctness o…

Model Selectionreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Predictive Complexity Priors

2020-06-18 · Eric Nalisnick, Jonathan Gordon, José Miguel Hernández-Lobato

Specifying a Bayesian prior is notoriously difficult for complex models such as neural networks. Reasoning about parameters is made challenging by the high-dimensionality and over-parameterization of the space. Priors th…

Few-Shot Learning

Deep Reference Priors: What is the best way to pretrain a model?

2022-02-01 · pproximateinference AABI Symposium 2022 2 · Yansong Gao, Rahul Ramesh, Pratik Chaudhari

What is the best way to exploit extra data -- be it unlabeled data from the same task, or labeled data from a related task -- to learn a given task? This paper formalizes the question using the theory of reference priors…

image-classificationSemi-Supervised Image ClassificationTransfer Learning

Comparison between Suitable Priors for Additive Bayesian Networks

2018-09-18 · Gilles Kratzer, Reinhard Furrer, Marta Pittavino

Additive Bayesian networks are types of graphical models that extend the usual Bayesian generalized linear model to multiple dependent variables through the factorisation of the joint probability distribution of the unde…

Model Selection