CVA Hedging by Risk-Averse Stochastic-Horizon Reinforcement Learning
This work studies the dynamic risk management of the risk-neutral value of the potential credit losses on a portfolio of derivatives. Sensitivities-based hedging of such liability is sub-optimal because of bid-ask costs, pricing models which cannot be completely realistic, and a discontinuity at default time. We leverage recent advances on risk-averse Reinforcement Learning developed specifically for option hedging with an ad hoc practice-aligned objective function aware of pathwise volatility, generalizing them to stochastic horizons. We formalize accurately the evolution of the hedger's portfolio stressing such aspects. We showcase the efficacy of our approach by a numerical study for a portfolio composed of a single FX forward contract.
Code (0)
등록된 구현이 없습니다.
Tasks
Managementreinforcement-learningReinforcement LearningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Deep Hedging: Continuous Reinforcement Learning for Hedging of General Portfolios across Multiple Risk Aversions
We present a method for finding optimal hedging policies for arbitrary initial portfolios and market states. We develop a novel actor-critic algorithm for solving general risk-averse stochastic control problems and use i…
Reinforcement Learning (RL)Option Hedging with Risk Averse Reinforcement Learning
In this paper we show how risk-averse reinforcement learning can be used to hedge options. We apply a state-of-the-art risk-averse algorithm: Trust Region Volatility Optimization (TRVO) to a vanilla option hedging enviro…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)A risk measurement approach from risk-averse stochastic optimization of score functions
We propose a risk measurement approach for a risk-averse stochastic problem. We provide results that guarantee that our problem has a solution. We characterize and explore the properties of the argmin as a risk measure a…
regressionStochastic OptimizationReinforcement Learning with Markov Risk Measures and Multipattern Risk Approximation
For a risk-averse finite-horizon Markov Decision Problem, we introduce a special class of Markov coherent risk measures, called mini-batch measures. We also define the class of multipattern risk-averse problems that gene…
Reinforcement LearningRobust Risk-Aware Option Hedging
The objectives of option hedging/trading extend beyond mere protection against downside risks, with a desire to seek gains also driving agent's strategies. In this study, we showcase the potential of robust risk-aware re…
Reinforcement Learning (RL)