Objective Variables for Probabilistic Revenue Maximization in Second-Price Auctions with Reserve
Many online companies sell advertisement space in second-price auctions with reserve. In this paper, we develop a probabilistic method to learn a profitable strategy to set the reserve price. We use historical auction data with features to fit a predictor of the best reserve price. This problem is delicate - the structure of the auction is such that a reserve price set too high is much worse than a reserve price set too low. To address this we develop objective variables, a new framework for combining probabilistic modeling with optimal decision-making. Objective variables are "hallucinated observations" that transform the revenue maximization task into a regularized maximum likelihood estimation problem, which we solve with an EM algorithm. This framework enables a variety of prediction mechanisms to set the reserve price. As examples, we study objective variable methods with regression, kernelized regression, and neural networks on simulated and real data. Our methods outperform previous approaches both in terms of scalability and profit.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingregressionSimilar Papers 제목 키워드 기반
Balancing Immediate Revenue and Future Off-Policy Evaluation in Coupon Allocation
Coupon allocation drives customer purchases and boosts revenue. However, it presents a fundamental trade-off between exploiting the current optimal policy to maximize immediate revenue and exploring alternative policies …
Off-policy evaluationRandomized Truthful Auctions with Learning Agents
We study a setting where agents use no-regret learning algorithms to participate in repeated auctions. \citet{kolumbus2022auctions} showed, rather surprisingly, that when bidders participate in second-price auctions usin…
Optimizing Revenue Maximization and Demand Learning in Airline Revenue Management
Correctly estimating how demand respond to prices is fundamental for airlines willing to optimize their pricing policy. Under some conditions, these policies, while aiming at maximizing short term revenue, can present to…
Demand ForecastingManagementDeep Reinforcement Learning Algorithm for Dynamic Pricing of Express Lanes with Multiple Access Locations
This article develops a deep reinforcement learning (Deep-RL) framework for dynamic pricing on managed lanes with multiple access locations and heterogeneity in travelers' value of time, origin, and destination. This fra…
Deep Reinforcement LearningPolicy Gradient MethodsReinforcement LearningReinforcement Learning (RL)Network Revenue Management with Demand Learning and Fair Resource-Consumption Balancing
In addition to maximizing the total revenue, decision-makers in lots of industries would like to guarantee balanced consumption across different resources. For instance, in the retailing industry, ensuring a balanced con…
Cloud ComputingFairnessManagement