paper-with-me

Papers

Stochastic Network Utility Maximization with Unknown Utilities: Multi-Armed Bandits Approach

2020-06-17 · Arun Verma, Manjesh K. Hanawal

In this paper, we study a novel Stochastic Network Utility Maximization (NUM) problem where the utilities of agents are unknown. The utility of each agent depends on the amount of resource it receives from a network operator/controller. The operator desires to do a resource allocation that maximizes the expected total utility of the network. We consider threshold type utility functions where each agent gets non-zero utility if the amount of resource it receives is higher than a certain threshold. Otherwise, its utility is zero (hard real-time). We pose this NUM setup with unknown utilities as a regret minimization problem. Our goal is to identify a policy that performs as `good' as an oracle policy that knows the utilities of agents. We model this problem setting as a bandit setting where feedback obtained in each round depends on the resource allocated to the agents. We propose algorithms for this novel setting using ideas from Multiple-Play Multi-Armed Bandits and Combinatorial Semi-Bandits. We show that the proposed algorithm is optimal when all agents have the same utility. We validate the performance guarantees of our proposed algorithms through numerical experiments.

📄 PDF Abstract BibTeX arXiv:2006.09997

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-Armed Bandits

Similar Papers 제목 키워드 기반

Stochastic Dynamic Network Utility Maximization with Application to Disaster Response

2024-06-06 · Anna Scaglione, Nurullah Karakoc

In this paper, we are interested in solving Network Utility Maximization (NUM) problems whose underlying local utilities and constraints depend on a complex stochastic dynamic environment. While the general model applies…

Deep Reinforcement LearningDisaster Response

Network Utility Maximization with Unknown Utility Functions: A Distributed, Data-Driven Bilevel Optimization Approach

2023-01-04 · Kaiyi Ji, Lei Ying

Fair resource allocation is one of the most important topics in communication networks. Existing solutions almost exclusively assume each user utility function is known and concave. This paper seeks to answer the followi…

Bilevel Optimization

Concave Statistical Utility Maximization Bandits via Influence-Function Gradients

2026-04-24 · Matías Carrasco, Alejandro Cholaquidis arxiv

We study stochastic multi-armed bandits in which the objective is a statistical functional of the long-run reward distribution, rather than expected reward alone. Under mild continuity assumptions, we show that the infin…

Multi-Armed Bandits

Stochastic Dominance Constrained Optimization with S-shaped Utilities: Poor-Performance-Region Algorithm and Neural Network

2025-11-29 · Zeyun Hu, Yang Liu arxiv

We investigate the static portfolio selection problem of S-shaped and non-concave utility maximization under first-order and second-order stochastic dominance (SD) constraints. In many S-shaped utility optimization probl…

Robust utility maximization with nonlinear continuous semimartingales

2022-06-28 · David Criens, Lars Niemann

In this paper we study a robust utility maximization problem in continuous time under model uncertainty. The model uncertainty is governed by a continuous semimartingale with uncertain local characteristics. Here, the di…