paper-with-me

홈 › Papers

Fast Distributed Bandits for Online Recommendation Systems

2020-07-16 · Kanak Mahadik, Qingyun Wu, Shuai Li, Amit Sabne

Contextual bandit algorithms are commonly used in recommender systems, where content popularity can change rapidly. These algorithms continuously learn latent mappings between users and items, based on contexts associated with them both. Recent recommendation algorithms that learn clustering or social structures between users have exhibited higher recommendation accuracy. However, as the number of users and items in the environment increases, the time required to generate recommendations deteriorates significantly. As a result, these cannot be deployed in practice. The state-of-the-art distributed bandit algorithm - DCCB - relies on a peer-to-peer net-work to share information among distributed workers. However, this approach does not scale well with the increasing number of users. Furthermore, it suffers from slow discovery of clusters, resulting in accuracy degradation. To address the above issues, this paper proposes a novel distributed bandit-based algorithm called DistCLUB. This algorithm lazily creates clusters in a distributed manner, and dramatically reduces the network data sharing requirement, achieving high scalability. Additionally, DistCLUB finds clusters much faster, achieving better accuracy than the state-of-the-art algorithm. Evaluation over both real-world benchmarks and synthetic datasets shows that DistCLUB is on average 8.87x faster than DCCB, and achieves 14.5% higher normalized prediction performance.

📄 PDF Abstract BibTeX arXiv:2007.08061

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringRecommendation Systems

Similar Papers 제목 키워드 기반

Preferences Evolve And So Should Your Bandits: Bandits with Evolving States for Online Platforms

2023-07-21 · Khashayar Khosravi, Renato Paes Leme, Chara Podimata, Apostolis Tsorvantzis

We propose a model for learning with bandit feedback while accounting for deterministically evolving and unobservable states that we call Bandits with Deterministically Evolving States ($B$-$DES$). The workhorse applicat…

Multi-Armed BanditsRecommendation Systems

Distributed Online Learning via Cooperative Contextual Bandits

2013-08-21 · Cem Tekin, Mihaela van der Schaar

In this paper we propose a novel framework for decentralized, online learning by many learners. At each moment of time, an instance characterized by a certain context may arrive to each learner; based on the context, the…

Event DetectionMulti-Armed BanditsRecommendation Systems

Online Matching: A Real-time Bandit System for Large-scale Recommendations

2023-07-29 · Xinyang Yi, Shao-Chuan Wang, Ruining He, Hariharan Chandrasekaran 외

The last decade has witnessed many successes of deep learning-based models for industry-scale recommender systems. These models are typically trained offline in a batch manner. While being effective in capturing users' p…

Multi-Armed BanditsRecommendation Systems

Neural Contextual Bandits for Personalized Recommendation

2023-12-21 · Yikun Ban, Yunzhe Qi, Jingrui He

In the dynamic landscape of online businesses, recommender systems are pivotal in enhancing user experiences. While traditional approaches have relied on static supervised learning, the quest for adaptive, user-centric r…

Multi-Armed BanditsRecommendation Systems

Neural Contextual Bandits Under Delayed Feedback Constraints

2025-04-16 · Mohammadali Moghimi, Sharu Theresa Jose, Shana Moothedath

This paper presents a new algorithm for neural contextual bandits (CBs) that addresses the challenge of delayed reward feedback, where the reward for a chosen action is revealed after a random, unknown delay. This scenar…

Multi-Armed BanditsRecommendation SystemsThompson Sampling