paper-with-me

Papers

Portfolio Reinforcement Learning with Scenario-Context Rollout

2026-02-27 · Vanya Priscillia Bendatu, Yao Lu arxiv

Market regime shifts induce distribution shifts that can degrade the performance of portfolio rebalancing policies. We propose macro-conditioned scenario-context rollout (SCR) that generates plausible next-day multivariate return scenarios under stress events. However, doing so faces new challenges, as history will never tell what would have happened differently. As a result, incorporating scenario-based rewards from rollouts introduces a reward--transition mismatch in temporal-difference learning, destabilizing RL critic training. We analyze this inconsistency and show it leads to a mixed evaluation target. Guided by this analysis, we construct a counterfactual next state using the rollout-implied continuations and augment the critic agent's bootstrap target. Doing so stabilizes the learning and provides a viable bias-variance tradeoff. In out-of-sample evaluations across 31 distinct universes of U.S. equity and ETF portfolios, our method improves Sharpe ratio by up to 76% and reduces maximum drawdown by up to 53% compared with classic and RL-based portfolio rebalancing baselines.

📄 PDF Abstract BibTeX arXiv:2602.24037

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

GIFT: LLM-Guided State-Reward Interface for Financial Reinforcement Learning

2026-06-07 · Yanyan Wu, Boyi Zhang, Yanlin Liu, Xinyu Fang 외 arxiv

Financial portfolio trading is naturally formulated as a reinforcement learning problem, where an agent sequentially rebalances assets under changing market conditions to balance return, risk, and transaction costs. Yet …

Reinforcement Learning

Adaptive Correlated Monte Carlo for Contextual Categorical Sequence Generation

2019-12-31 · ICLR 2020 1 · Xinjie Fan, Yizhe Zhang, Zhendong Wang, Mingyuan Zhou

Sequence generation models are commonly refined with reinforcement learning over user-defined metrics. However, high gradient variance hinders the practical use of this method. To stabilize this method, we adapt to conte…

Image CaptioningProgram SynthesisReinforcement Learning

BubbleSpec: Turning Long-Tail Bubbles into Speculative Rollout Drafts for Synchronous Reinforcement Learning

2026-05-09 · Yuhang Xu, Kaibin Tian, Yang Tian, Zhice Yang 외 arxiv

Reinforcement Learning (RL) has become a cornerstone for improving the performance of Large Language Models (LLMs). However, its rollout phase constitutes a significant efficiency bottleneck, mainly arising from the long…

Reinforcement Learning

Markowitz Meets Bellman: Knowledge-distilled Reinforcement Learning for Portfolio Management

2024-05-08 · Gang Hu, Ming Gu

Investment portfolios, central to finance, balance potential returns and risks. This paper introduces a hybrid approach combining Markowitz's portfolio theory with reinforcement learning, utilizing knowledge distillation…

Knowledge DistillationManagementreinforcement-learningReinforcement Learning

The Exploratory Multi-Asset Mean-Variance Portfolio Selection using Reinforcement Learning

2025-05-12 · Yu Li, Yuhan Wu, Shuhua Zhang

In this paper, we study the continuous-time multi-asset mean-variance (MV) portfolio selection using a reinforcement learning (RL) algorithm, specifically the soft actor-critic (SAC) algorithm, in the time-varying financ…

Reinforcement Learning (RL)