paper-with-me

Papers

Characterizing Truthful Multi-Armed Bandit Mechanisms

2008-12-12 · Moshe Babaioff, Yogeshwer Sharma, Aleksandrs Slivkins

We consider a multi-round auction setting motivated by pay-per-click auctions for Internet advertising. In each round the auctioneer selects an advertiser and shows her ad, which is then either clicked or not. An advertiser derives value from clicks; the value of a click is her private information. Initially, neither the auctioneer nor the advertisers have any information about the likelihood of clicks on the advertisements. The auctioneer's goal is to design a (dominant strategies) truthful mechanism that (approximately) maximizes the social welfare. If the advertisers bid their true private values, our problem is equivalent to the "multi-armed bandit problem", and thus can be viewed as a strategic version of the latter. In particular, for both problems the quality of an algorithm can be characterized by "regret", the difference in social welfare between the algorithm and the benchmark which always selects the same "best" advertisement. We investigate how the design of multi-armed bandit algorithms is affected by the restriction that the resulting mechanism must be truthful. We find that truthful mechanisms have certain strong structural properties -- essentially, they must separate exploration from exploitation -- and they incur much higher regret than the optimal multi-armed bandit algorithms. Moreover, we provide a truthful mechanism which (essentially) matches our lower bound on regret.

📄 PDF Abstract BibTeX arXiv:0812.2291

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the Robustness of Epoch-Greedy in Multi-Agent Contextual Bandit Mechanisms

2023-07-15 · Yinglun Xu, Bhuvesh Kumar, Jacob Abernethy

Efficient learning in multi-armed bandit mechanisms such as pay-per-click (PPC) auctions typically involves three challenges: 1) inducing truthful bidding behavior (incentives), 2) using personalization in the users (con…

Auction-Based Combinatorial Multi-Armed Bandit Mechanisms with Strategic Arms

2021-05-10 · IEEE Conference on Computer Communications 2021 5 · Guoju Gao, He Huang, Mingjun Xiao, Jie Wu 외

The multi-armed bandit (MAB) model has been deeply studied to solve many online learning problems, such as rate allocation in communication networks, Ad recommendation in social networks, etc. In an MAB model, given N ar…

Computational Efficiency

Designing Truthful Contextual Multi-Armed Bandits based Sponsored Search Auctions

2020-02-26 · Kumar Abhishek, Shweta Jain, Sujit Gujar

For sponsored search auctions, we consider contextual multi-armed bandit problem in the presence of strategic agents. In this setting, at each round, an advertising platform (center) runs an auction to select the best-su…

Multi-Armed Bandits

Strategic Multi-Armed Bandit Problems Under Debt-Free Reporting

2025-01-27 · Ahmed Ben Yahmed, Clément Calauzènes, Vianney Perchet

We consider the classical multi-armed bandit problem, but with strategic arms. In this context, each arm is characterized by a bounded support reward distribution and strategically aims to maximize its own utility by pot…

Simple regret for infinitely many armed bandits

2015-05-18 · Alexandra Carpentier, Michal Valko

We consider a stochastic bandit problem with infinitely many arms. In this setting, the learner has no chance of trying all the arms even once and has to dedicate its limited number of samples only to a certain number of…