paper-with-me

홈 › Papers

Context Uncertainty in Contextual Bandits with Applications to Recommender Systems

2022-02-01 · Hao Wang, Yifei Ma, Hao Ding, Yuyang Wang

Recurrent neural networks have proven effective in modeling sequential user feedbacks for recommender systems. However, they usually focus solely on item relevance and fail to effectively explore diverse items for users, therefore harming the system performance in the long run. To address this problem, we propose a new type of recurrent neural networks, dubbed recurrent exploration networks (REN), to jointly perform representation learning and effective exploration in the latent space. REN tries to balance relevance and exploration while taking into account the uncertainty in the representations. Our theoretical analysis shows that REN can preserve the rate-optimal sublinear regret even when there exists uncertainty in the learned representations. Our empirical study demonstrates that REN can achieve satisfactory long-term rewards on both synthetic and real-world recommendation datasets, outperforming state-of-the-art models.

📄 PDF Abstract BibTeX arXiv:2202.00805

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-Armed BanditsRecommendation SystemsRepresentation Learning

Similar Papers 제목 키워드 기반

Neural Contextual Bandits for Personalized Recommendation

2023-12-21 · Yikun Ban, Yunzhe Qi, Jingrui He

In the dynamic landscape of online businesses, recommender systems are pivotal in enhancing user experiences. While traditional approaches have relied on static supervised learning, the quest for adaptive, user-centric r…

Multi-Armed BanditsRecommendation Systems

$α$-Fair Contextual Bandits

2023-10-22 · Siddhant Chaudhary, Abhishek Sinha

Contextual bandit algorithms are at the core of many applications, including recommender systems, clinical trials, and optimal portfolio selection. One of the most popular problems studied in the contextual bandit litera…

Multi-Armed BanditsRecommendation Systems

Deep Contextual Multi-armed Bandits

2018-07-25 · Mark Collier, Hector Urdiales Llorens

Contextual multi-armed bandit problems arise frequently in important industrial applications. Existing solutions model the context either linearly, which enables uncertainty driven (principled) exploration, or non-linear…

MarketingMulti-Armed BanditsThompson Sampling

Neural Contextual Bandits without Regret

2021-07-07 · Parnian Kassraie, Andreas Krause

Contextual bandits are a rich model for sequential decision making given side information, with important applications, e.g., in recommender systems. We propose novel algorithms for contextual bandits harnessing neural n…

Decision MakingMulti-Armed BanditsRecommendation SystemsSequential Decision Making

Syndicated Bandits: A Framework for Auto Tuning Hyper-parameters in Contextual Bandit Algorithms

2021-06-05 · Qin Ding, Yue Kang, Yi-Wei Liu, Thomas C. M. Lee 외

The stochastic contextual bandit problem, which models the trade-off between exploration and exploitation, has many real applications, including recommender systems, online advertising and clinical trials. As many other …

Recommendation Systems