paper-with-me

홈 › Papers

Joint Online Learning and Decision-making via Dual Mirror Descent

2021-04-20 · Alfonso Lobos, Paul Grigas, Zheng Wen

We consider an online revenue maximization problem over a finite time horizon subject to lower and upper bounds on cost. At each period, an agent receives a context vector sampled i.i.d. from an unknown distribution and needs to make a decision adaptively. The revenue and cost functions depend on the context vector as well as some fixed but possibly unknown parameter vector to be learned. We propose a novel offline benchmark and a new algorithm that mixes an online dual mirror descent scheme with a generic parameter learning process. When the parameter vector is known, we demonstrate an $O(\sqrt{T})$ regret result as well an $O(\sqrt{T})$ bound on the possible constraint violations. When the parameter is not known and must be learned, we demonstrate that the regret and constraint violations are the sums of the previous $O(\sqrt{T})$ terms plus terms that directly depend on the convergence of the learning process.

📄 PDF Abstract BibTeX arXiv:2104.09750

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Online Contextual Decision-Making with a Smart Predict-then-Optimize Method

2022-06-15 · Heyuan Liu, Paul Grigas

We study an online contextual decision-making problem with resource constraints. At each time period, the decision-maker first predicts a reward vector and resource consumption matrix based on a given context vector and …

Decision MakingPrediction

Online Sequential Decision-Making with Unknown Delays

2024-02-12 · Ping Wu, Heyan Huang, Zhengyang Liu

In the field of online sequential decision-making, we address the problem with delays utilizing the framework of online convex optimization (OCO), where the feedback of a decision can arrive with an unknown delay. Unlike…

Decision MakingSequential Decision Making

Online Algorithmic Recourse by Collective Action

2023-12-29 · Elliot Creager, Richard Zemel

Research on algorithmic recourse typically considers how an individual can reasonably change an unfavorable automated decision when interacting with a fixed decision-making system. This paper focuses instead on the onlin…

Decision Making

The Best of Many Worlds: Dual Mirror Descent for Online Allocation Problems

2020-11-18 · Santiago Balseiro, Haihao Lu, Vahab Mirrokni

Online allocation problems with resource constraints are central problems in revenue management and online advertising. In these problems, requests arrive sequentially during a finite horizon and, for each request, a dec…

Assortment OptimizationManagement

To Mask or to Mirror: Human-AI Alignment in Collective Reasoning

2025-10-02 · Crystal Qian, Aaron Parisi, Clémentine Bouleau, Vivian Tsai 외 arxiv

As large language models (LLMs) are increasingly used to model and augment collective decision-making, it is critical to examine their alignment with human social reasoning. We present an empirical framework for assessin…