paper-with-me

Papers

Risk-averse learning with delayed feedback

2024-09-25 · Siyi Wang, Zifan Wang, Karl Henrik Johansson, Sandra Hirche

In real-world scenarios, the impacts of decisions may not manifest immediately. Taking these delays into account facilitates accurate assessment and management of risk in real-world environments, thereby ensuring the efficacy of strategies. In this paper, we investigate risk-averse learning using Conditional Value at Risk (CVaR) as risk measure, while incorporating delayed feedback with unknown but bounded delays. We develop two risk-averse learning algorithms that rely on one-point and two-point zeroth-order optimization approaches, respectively. The regret achieved by the algorithms is analyzed in terms of the cumulative delay and the number of total samplings. The results suggest that the two-point risk-averse learning achieves a smaller regret bound than the one-point algorithm. Furthermore, the one-point risk-averse learning algorithm attains sublinear regret under certain delay conditions, and the two-point risk-averse learning algorithm achieves sublinear regret with minimal restrictions on the delay. We provide numerical experiments on a dynamic pricing problem to demonstrate the performance of the proposed algorithms.

📄 PDF Abstract BibTeX arXiv:2409.16866

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

A Zeroth-Order Momentum Method for Risk-Averse Online Convex Games

2022-09-06 · Zifan Wang, Yi Shen, Zachary I. Bell, Scott Nivison 외

We consider risk-averse learning in repeated unknown games where the goal of the agents is to minimize their individual risk of incurring significantly high cost. Specifically, the agents use the conditional value at ris…

Risk-Averse Finetuning of Large Language Models

2025-01-12 · Sapana Chaudhary, Ujwal Dinesha, Dileep Kalathil, Srinivas Shakkottai

We consider the challenge of mitigating the generation of negative or toxic content by the Large Language Models (LLMs) in response to certain prompts. We propose integrating risk-averse principles into LLM fine-tuning t…

Risk-Averse Stochastic Convex Bandit

2018-10-01 · Adrian Rivera Cardoso, Huan Xu

Motivated by applications in clinical trials and finance, we study the problem of online convex optimization (with bandit feedback) where the decision maker is risk-averse. We provide two algorithms to solve this problem…

Risk-Averse No-Regret Learning in Online Convex Games

2022-03-16 · Zifan Wang, Yi Shen, Michael M. Zavlanos

We consider an online stochastic game with risk-averse agents whose goal is to learn optimal decisions that minimize the risk of incurring significantly high costs. Specifically, we use the Conditional Value at Risk (CVa…

Risk-Averse Receding Horizon Motion Planning for Obstacle Avoidance using Coherent Risk Measures

2022-04-20 · Anushri Dixit, Mohamadreza Ahmadi, Joel W. Burdick

This paper studies the problem of risk-averse receding horizon motion planning for agents with uncertain dynamics, in the presence of stochastic, dynamic obstacles. We propose a model predictive control (MPC) scheme that…

Model Predictive ControlMotion Planning