paper-with-me

Papers

Auction-Based Combinatorial Multi-Armed Bandit Mechanisms with Strategic Arms

2021-05-10 · IEEE Conference on Computer Communications 2021 5 · Guoju Gao, He Huang, Mingjun Xiao, Jie Wu, Yu-E Sun, Sheng Zhang

The multi-armed bandit (MAB) model has been deeply studied to solve many online learning problems, such as rate allocation in communication networks, Ad recommendation in social networks, etc. In an MAB model, given N arms whose rewards are unknown in advance, the player selects exactly one arm in each round, and his goal is to maximize the cumulative rewards over a fixed horizon. In this paper, we study the budget-constrained auction-based combinatorial multi-armed bandit mechanism with strategic arms, where the player can select K (<; N) arms in a round and pulling each arm has a unique cost. In addition, each arm might strategically report its cost in the auction. To this end, we combine the upper confidence bound (UCB) with auction to define the UCB-based rewards and then devise an auction-based UCB algorithm (called AUCB). In each round, AUCB selects the top K arms according to the ratios of UCB-based rewards to bids and further determines the critical payment for each arm. For AUCB, we derive an upper bound on regret and prove the truthfulness, individual rationality, and computational efficiency. Extensive simulations show that the rewards achieved by AUCB are at least 12.49% higher than those of state-of-the-art algorithms.

📄 PDF Abstract BibTeX

Code (1)

DURUII/Replica-AUCB

Tasks

Computational Efficiency

Similar Papers 제목 키워드 기반

Characterizing Truthful Multi-Armed Bandit Mechanisms

2008-12-12 · Moshe Babaioff, Yogeshwer Sharma, Aleksandrs Slivkins

We consider a multi-round auction setting motivated by pay-per-click auctions for Internet advertising. In each round the auctioneer selects an advertiser and shows her ad, which is then either clicked or not. An adverti…

A Semi-supervised Sensing Rate Learning based CMAB Scheme to Combat COVID-19 by Trustful Data Collection in the Crowd

2023-01-17 · Jianheng Tang, Kejia Fan, Wenxuan Xie, Luomin Zeng 외

The recruitment of trustworthy and high-quality workers is an important research issue for MCS. Previous studies either assume that the qualities of workers are known in advance, or assume that the platform knows the qua…

On the Robustness of Epoch-Greedy in Multi-Agent Contextual Bandit Mechanisms

2023-07-15 · Yinglun Xu, Bhuvesh Kumar, Jacob Abernethy

Efficient learning in multi-armed bandit mechanisms such as pay-per-click (PPC) auctions typically involves three challenges: 1) inducing truthful bidding behavior (incentives), 2) using personalization in the users (con…

Designing Truthful Contextual Multi-Armed Bandits based Sponsored Search Auctions

2020-02-26 · Kumar Abhishek, Shweta Jain, Sujit Gujar

For sponsored search auctions, we consider contextual multi-armed bandit problem in the presence of strategic agents. In this setting, at each round, an advertising platform (center) runs an auction to select the best-su…

Multi-Armed Bandits

Online Causal Inference for Advertising in Real-Time Bidding Auctions

2019-08-22 · Caio Waisman, Harikesh S. Nair, Carlos Carrion

Real-time bidding (RTB) systems, which utilize auctions to allocate user impressions to competing advertisers, continue to enjoy success in digital advertising. Assessing the effectiveness of such advertising remains a c…

Causal InferenceExperimental DesignThompson Sampling