paper-with-me

홈 › Papers

Functional Bandits

2014-05-10 · Long Tran-Thanh, Jia Yuan Yu

We introduce the functional bandit problem, where the objective is to find an arm that optimises a known functional of the unknown arm-reward distributions. These problems arise in many settings such as maximum entropy methods in natural language processing, and risk-averse decision-making, but current best-arm identification techniques fail in these domains. We propose a new approach, that combines functional estimation and arm elimination, to tackle this problem. This method achieves provably efficient performance guarantees. In addition, we illustrate this method on a number of important functionals in risk management and information theory, and refine our generic theoretical results in those cases.

📄 PDF Abstract BibTeX arXiv:1405.2432

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingManagement

Similar Papers 제목 키워드 기반

Efficient Contextual Bandits with Continuous Actions

2020-06-10 · NeurIPS 2020 12 · Maryam Majzoubi, Chicheng Zhang, Rajan Chari, Akshay Krishnamurthy 외

We create a computationally tractable algorithm for contextual bandits with continuous actions having unknown structure. Our reduction-style algorithm composes with most supervised learning representations. We prove that…

Multi-Armed Bandits

Signature Approach for Contextual Bandits with Nonlinear and Path-dependent Rewards

2026-05-11 · Xin Guo, Grace He, Xinyu Li arxiv

We study contextual bandits with nonlinear and path-dependent rewards through a novel signature-transform-based approach. Leveraging the universal nonlinearity property of signatures, we approximate continuous path-depen…

Off-Policy Risk Assessment in Contextual Bandits

2021-04-18 · NeurIPS 2021 12 · Audrey Huang, Liu Leqi, Zachary C. Lipton, Kamyar Azizzadenesheli

Even when unable to run experiments, practitioners can evaluate prospective policies, using previously logged data. However, while the bandits literature has adopted a diverse set of objectives, most research on off-poli…

Multi-Armed BanditsOff-policy evaluation

Contextual Online Decision Making with Infinite-Dimensional Functional Regression

2025-01-30 · Haichen Hu, Rui Ai, Stephen Bates, David Simchi-Levi

Contextual sequential decision-making problems play a crucial role in machine learning, encompassing a wide range of downstream applications such as bandits, sequential hypothesis testing and online risk control. These a…

Decision MakingMulti-Armed BanditsregressionSequential Decision Making

A Deep Bayesian Bandits Approach for Anticancer Therapy: Exploration via Functional Prior

2022-05-05 · Mingyu Lu, Yifang Chen, Su-In Lee

Learning personalized cancer treatment with machine learning holds great promise to improve cancer patients' chance of survival. Despite recent advances in machine learning and precision oncology, this approach remains c…

BIG-bench Machine LearningDrug Response Prediction