paper-with-me

홈 › Papers

Learning optimal Bayesian prior probabilities from data

2021-01-03 · Ozan Kaan Kayaalp

Noninformative uniform priors are staples of Bayesian inference, especially in Bayesian machine learning. This study challenges the assumption that they are optimal and their use in Bayesian inference yields optimal outcomes. Instead of using arbitrary noninformative uniform priors, we propose a machine learning based alternative method, learning optimal priors from data by maximizing a target function of interest. Applying na\"ive Bayes text classification methodology and a search algorithm developed for this study, our system learned priors from data using the positive predictive value metric as the target function. The task was to find Wikipedia articles that had not (but should have) been categorized under certain Wikipedia categories. We conducted five sets of experiments using separate Wikipedia categories. While the baseline models used the popular Bayes-Laplace priors, the study models learned the optimal priors for each set of experiments separately before using them. The results showed that the study models consistently outperformed the baseline models with a wide margin of statistical significance (p < 0.001). The measured performance improvement of the study model over the baseline was as high as 443% with the mean value of 193% over five Wikipedia categories.

📄 PDF Abstract BibTeX arXiv:2101.00672

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesBayesian InferenceBIG-bench Machine Learningtext-classificationText Classification

Similar Papers 제목 키워드 기반

Maximum Entropy competes with Maximum Likelihood

2020-12-17 · A. E. Allahverdyan, N. H. Martirosyan

Maximum entropy (MAXENT) method has a large number of applications in theoretical and applied machine learning, since it provides a convenient non-parametric tool for estimating unknown probabilities. The method is a maj…

Maximum Entropy competes with Maximum Likelihood

2020-09-28 · Armen Allahverdyan

Maximum entropy (MAXENT) method has a large number of applications in theoretical and applied machine learning, since it provides a convenient non-parametric tool for estimating unknown probabilities. The method is a ma…

The Bayesian Origin of the Probability Weighting Function in Human Representation of Probabilities

2025-10-06 · Xin Tong, Thi Thu Uyen Hoang, Xue-Xin Wei, Michael Hahn arxiv

Humans systematically misrepresent probability in a stereotyped inverse-S pattern. It has been documented for decades, but its origin remains unexplained. We propose a Bayesian encoding-decoding account in which probabil…

Beyond Uniform Priors in Bayesian Network Structure Learning

2017-04-12 · Marco Scutari

Bayesian network structure learning is often performed in a Bayesian setting, evaluating candidate structures using their posterior probabilities for a given data set. Score-based algorithms then use those posterior prob…

Big Batch Bayesian Active Learning by Considering Predictive Probabilities

2025-01-14 · Sebastian W. Ober, Samuel Power, Tom Diethe, Henry B. Moss

We observe that BatchBALD, a popular acquisition function for batch Bayesian active learning for classification, can conflate epistemic and aleatoric uncertainty, leading to suboptimal performance. Motivated by this obse…

Active Learning