paper-with-me

Papers

Learning to Bid Without Knowing your Value

2017-11-03 · Zhe Feng, Chara Podimata, Vasilis Syrgkanis

We address online learning in complex auction settings, such as sponsored search auctions, where the value of the bidder is unknown to her, evolving in an arbitrary manner and observed only if the bidder wins an allocation. We leverage the structure of the utility of the bidder and the partial feedback that bidders typically receive in auctions, in order to provide algorithms with regret rates against the best fixed bid in hindsight, that are exponentially faster in convergence in terms of dependence on the action space, than what would have been derived by applying a generic bandit algorithm and almost equivalent to what would have been achieved in the full information setting. Our results are enabled by analyzing a new online learning setting with outcome-based feedback, which generalizes learning with feedback graphs. We provide an online learning algorithm for this setting, of independent interest, with regret that grows only logarithmically with the number of actions and linearly only in the number of potential outcomes (the latter being very small in most auction settings). Last but not least, we show that our algorithm outperforms the bandit approach experimentally and that this performance is robust to dropping some of our theoretical assumptions or introducing noise in the feedback that the bidder receives.

📄 PDF Abstract BibTeX arXiv:1711.01333

Code (1)

zfengharvard/bandit-sponsored-search 공식 구현

Similar Papers 제목 키워드 기반

"Knowing value" logic as a normal modal logic

2016-04-29 · Tao Gu, Yanjing Wang

Recent years witness a growing interest in nonstandard epistemic logics of "knowing whether", "knowing what", "knowing how", and so on. These logics are usually not normal, i.e., the standard axioms and reasoning rules f…

Negation

ModelLock: Locking Your Model With a Spell

2024-05-25 · Yifeng Gao, Yuhua Sun, Xingjun Ma, Zuxuan Wu 외

This paper presents a novel model protection paradigm ModelLock that locks (destroys) the performance of a model on normal clean data so as to make it unusable or unextractable without the right key. Specifically, we pro…

image-classificationImage Classificationmodeltext-guided-image-editing

The Value of Information in Human-AI Decision-making

2025-02-10 · Ziyang Guo, Yifan Wu, Jason Hartline, Jessica Hullman

Multiple agents -- including humans and AI models -- are often paired on decision tasks with the expectation of achieving complementary performance, where the combined performance of both agents outperforms either one al…

Decision MakingModel Selection

Identifying two piecewise linear additive value functions from anonymous preference information

2026-02-24 · Vincent Auriau, Khaled Belahcene, Emmanuel Malherbe, Vincent Mousseau 외 arxiv

Eliciting a preference model involves asking a person, named decision-maker, a series of questions. We assume that these preferences can be represented by an additive value function. In this work, we query simultaneously…

Knowing Whether

2013-11-30 · Jie Fan, Yanjing Wang, Hans van Ditmarsch

Knowing whether a proposition is true means knowing that it is true or knowing that it is false. In this paper, we study logics with a modal operator Kw for knowing whether but without a modal operator K for knowing that…