paper-with-me

홈 › Papers

Risk and Regret of Hierarchical Bayesian Learners

2015-05-19 · Jonathan H. Huggins, Joshua B. Tenenbaum

Common statistical practice has shown that the full power of Bayesian methods is not realized until hierarchical priors are used, as these allow for greater "robustness" and the ability to "share statistical strength." Yet it is an ongoing challenge to provide a learning-theoretically sound formalism of such notions that: offers practical guidance concerning when and how best to utilize hierarchical models; provides insights into what makes for a good hierarchical prior; and, when the form of the prior has been chosen, can guide the choice of hyperparameter settings. We present a set of analytical tools for understanding hierarchical priors in both the online and batch learning settings. We provide regret bounds under log-loss, which show how certain hierarchical models compare, in retrospect, to the best single model in the model class. We also show how to convert a Bayesian log-loss regret bound into a Bayesian risk bound for any bounded loss, a result which may be of independent interest. Risk and regret bounds for Student's $t$ and hierarchical Gaussian priors allow us to formalize the concepts of "robustness" and "sharing statistical strength." Priors for feature selection are investigated as well. Our results suggest that the learning-theoretic benefits of using hierarchical priors can often come at little cost on practical problems.

📄 PDF Abstract BibTeX arXiv:1505.04984

Code (0)

등록된 구현이 없습니다.

Tasks

feature selection

Similar Papers 제목 키워드 기반

Markets with Heterogeneous Agents: Dynamics and Survival of Bayesian vs. No-Regret Learners

2025-02-12 · David Easley, Yoav Kolumbus, Eva Tardos

We analyze the performance of heterogeneous learning agents in asset markets with stochastic payoffs. Our main focus is on comparing Bayesian learners and no-regret learners who compete in markets and identifying the con…

Learning Theory

Bayesian Online Model Selection

2026-02-20 · Aida Afshar, Yuke Zhang, Aldo Pacchiano arxiv

Online model selection in Bayesian bandits raises a fundamental exploration challenge: When an environment instance is sampled from a prior distribution, how can we design an adaptive strategy that explores multiple band…

Online Bayesian Risk-Averse Reinforcement Learning

2025-09-17 · Yuhao Wang, Enlu Zhou arxiv

In this paper, we study the Bayesian risk-averse formulation in reinforcement learning (RL). To address the epistemic uncertainty due to a lack of data, we adopt the Bayesian Risk Markov Decision Process (BRMDP) to accou…

Reinforcement Learning

Golden Handcuffs make safer AI agents

2026-04-15 · Aram Ebtekar, Michael K. Cohen arxiv

Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand the agent's subjective reward range to include a large negative value …

Strategizing against No-Regret Learners in First-Price Auctions

2024-02-13 · Aviad Rubinstein, Junyao Zhao

We study repeated first-price auctions and general repeated Bayesian games between two players, where one player, the learner, employs a no-regret learning algorithm, and the other player, the optimizer, knowing the lear…