paper-with-me

홈 › Papers

Learning to Price Against a Moving Target

2021-06-08 · Renato Paes Leme, Balasubramanian Sivan, Yifeng Teng, Pratik Worah

In the Learning to Price setting, a seller posts prices over time with the goal of maximizing revenue while learning the buyer's valuation. This problem is very well understood when values are stationary (fixed or iid). Here we study the problem where the buyer's value is a moving target, i.e., they change over time either by a stochastic process or adversarially with bounded variation. In either case, we provide matching upper and lower bounds on the optimal revenue loss. Since the target is moving, any information learned soon becomes out-dated, which forces the algorithms to keep switching between exploring and exploiting phases.

📄 PDF Abstract BibTeX arXiv:2106.04689

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Optimal Solutions for the Moving Target Vehicle Routing Problem via Branch-and-Price with Relaxed Continuity

2026-02-28 · Anoop Bhat, Geordan Gutow, Zhongqiang Ren, Sivakumar Rathinam 외 arxiv

The Moving Target Vehicle Routing Problem (MT-VRP) seeks trajectories for several agents that intercept a set of moving targets, subject to speed, time window, and capacity constraints. We introduce an exact algorithm, B…

Optimal Solutions for the Moving Target Vehicle Routing Problem with Obstacles via Lazy Branch and Price

2026-03-23 · Anoop Bhat, Geordan Gutow, Surya Singh, Zhongqiang Ren 외 arxiv

The Moving Target Vehicle Routing Problem with Obstacles (MT-VRP-O) seeks trajectories for several agents that collectively intercept a set of moving targets. Each target has one or more time windows where it must be vis…

Motion Planning

TT-DAC-PS: Twin-Target Deterministic Actor-Critic with Policy Smoothing for Optimal Trade Execution

2026-06-07 · Ilia Zaznov, Atta Badii, Julian Kunkel, Alfonso Dufour arxiv

This study addresses the optimal execution of large stock sell programs by introducing TT-DAC-PS (Twin-Target Deterministic Actor-Critic with Policy Smoothing), a deterministic actor-critic architecture that combines twi…

Improved Forecasting of Cryptocurrency Price using Social Signals

2019-07-01 · Maria Glenski, Tim Weninger, Svitlana Volkova

Social media signals have been successfully used to develop large-scale predictive and anticipatory analytics. For example, forecasting stock market prices and influenza outbreaks. Recently, social data has been explored…

Two-Phase Bilevel Search for the Moving-Target Traveling Salesman Problem with Moving Obstacles

2026-06-17 · Allen George Philip, Anoop Bhat, Sivakumar Rathinam, Howie Choset arxiv

The Moving-Target Traveling Salesman Problem (MT-TSP) seeks a minimum cost trajectory for an agent that departs from a static depot, visits a set of moving targets, each within one of their assigned time windows, and ret…