paper-with-me

Papers

Dynamic pricing and assortment under a contextual MNL demand

2021-10-19 · Vineet Goyal, Noemie Perivier

We consider dynamic multi-product pricing and assortment problems under an unknown demand over T periods, where in each period, the seller decides on the price for each product or the assortment of products to offer to a customer who chooses according to an unknown Multinomial Logit Model (MNL). Such problems arise in many applications, including online retail and advertising. We propose a randomized dynamic pricing policy based on a variant of the Online Newton Step algorithm (ONS) that achieves a $O(d\sqrt{T}\log(T))$ regret guarantee under an adversarial arrival model. We also present a new optimistic algorithm for the adversarial MNL contextual bandits problem, which achieves a better dependency than the state-of-the-art algorithms in a problem-dependent constant $\kappa_2$ (potentially exponentially small). Our regret upper bound scales as $\tilde{O}(d\sqrt{\kappa_2 T}+ \log(T)/\kappa_2)$, which gives a stronger bound than the existing $\tilde{O}(d\sqrt{T}/\kappa_2)$ guarantees.

📄 PDF Abstract BibTeX arXiv:2110.10018

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-Armed Bandits

Similar Papers 제목 키워드 기반

Poisson-MNL Bandit: Nearly Optimal Dynamic Joint Assortment and Pricing with Decision-Dependent Customer Arrivals

2026-02-18 · Junhui Cai, Ran Chen, Qitao Huang, Linda Zhao 외 arxiv

We study dynamic joint assortment and pricing where a seller updates decisions at regular accounting/operating intervals to maximize the cumulative per-period revenue over a horizon $T$. In many settings, assortment and …

Uncertainty Quantification for Demand Prediction in Contextual Dynamic Pricing

2020-03-16 · Yining Wang, Xi Chen, Xiangyu Chang, Dongdong Ge

Data-driven sequential decision has found a wide range of applications in modern operations management, such as dynamic pricing, inventory control, and assortment optimization. Most existing research on data-driven seque…

Assortment OptimizationManagementUncertainty Quantificationvalid

Doubly High-Dimensional Contextual Bandits: An Interpretable Model for Joint Assortment-Pricing

2023-09-14 · Junhui Cai, Ran Chen, Martin J. Wainwright, Linda Zhao

Key challenges in running a retail business include how to select products to present to consumers (the assortment problem), and how to price products (the pricing problem) to maximize revenue or profit. Instead of consi…

Multi-Armed Bandits

Transfer Learning for Contextual Joint Assortment-Pricing under Cross-Market Heterogeneity

2026-03-18 · Elynn Chen, Xi Chen, Yi Zhang arxiv

We study transfer learning for contextual joint assortment-pricing under a multinomial logit choice model with bandit feedback. A seller operates across multiple related markets and observes only posted prices and realiz…

Transfer Learning

Online Assortment and Price Optimization Under Contextual Choice Models

2025-03-14 · Yigit Efe Erginbas, Thomas A. Courtade, Kannan Ramchandran

We consider an assortment selection and pricing problem in which a seller has $N$ different items available for sale. In each round, the seller observes a $d$-dimensional contextual preference information vector for the …