paper-with-me

Papers

Time Consistent Stopping For The Mean-Standard Deviation Problem --- The Discrete Time Case

2019-04-19

Inspired by Strotz's consistent planning strategy, we formulate the infinite horizon mean-variance stopping problem as a subgame perfect Nash equilibrium in order to determine time consistent strategies with no regret. Equilibria among stopping times or randomized stopping times may not exist. This motivates us to consider the notion of liquidation strategies, which lets the stopping right to be divisible. We then argue that the mean-standard deviation variant of this problem makes more sense for this type of strategies in terms of time consistency. It turns out that an equilibrium liquidation strategy always exists. We then analyze whether optimal equilibrium liquidation strategies exist and whether they are unique and observe that neither may hold.

📄 PDF Abstract BibTeX arXiv:1802.08358

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Mixture Martingales Revisited with Applications to Sequential Tests and Confidence Intervals

2018-11-28 · Emilie Kaufmann, Wouter Koolen

This paper presents new deviation inequalities that are valid uniformly in time under adaptive sampling in a multi-armed bandit model. The deviations are measured using the Kullback-Leibler divergence in a given one-dime…

valid

Optimal Stopping Theory for a Distributionally Robust Seller

2022-06-06 · Pieter Kleer, Johan van Leeuwaarden

Sellers in online markets face the challenge of determining the right time to sell in view of uncertain future offers. Classical stopping theory assumes that sellers have full knowledge of the value distributions, and le…

Sharp Matrix Empirical Bernstein Inequalities

2024-11-14 · Hongjian Wang, Aaditya Ramdas

We present two sharp, closed-form empirical Bernstein inequalities for symmetric random matrices with bounded eigenvalues. By sharp, we mean that both inequalities adapt to the unknown variance in a tight manner: the dev…

Measures of Variability for Risk-averse Policy Gradient

2025-04-15 · Yudong Luo, Yangchen Pan, Jiaqi Tan, Pascal Poupart

Risk-averse reinforcement learning (RARL) is critical for decision-making under uncertainty, which is especially valuable in high-stake applications. However, most existing works focus on risk measures, e.g., conditional…

Decision MakingDecision Making Under Uncertainty

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade

2026-07-07 · Kai Ruan, Zihe Huang, Ziqi Zhou, Qianshan Wei 외 arxiv

Large language model (LLM) agents often waste inference compute by continuing multi-step trajectories that are already doomed to fail. We study early failure prediction and inference-time early stopping for LLM agents us…