paper-with-me

홈 › Papers

FR-LUX: Friction-Aware, Regime-Conditioned Policy Optimization for Implementable Portfolio Management

2025-10-03 · Jian'an Zhang arxiv

Transaction costs and regime shifts are major reasons why paper portfolios fail in live trading. We introduce FR-LUX (Friction-aware, Regime-conditioned Learning under eXecution costs), a reinforcement learning framework that learns after-cost trading policies and remains robust across volatility-liquidity regimes. FR-LUX integrates three ingredients: (i) a microstructure-consistent execution model combining proportional and impact costs, directly embedded in the reward; (ii) a trade-space trust region that constrains changes in inventory flow rather than logits, yielding stable low-turnover updates; and (iii) explicit regime conditioning so the policy specializes to LL/LH/HL/HH states without fragmenting the data. On a 4 x 5 grid of regimes and cost levels with multiple random seeds, FR-LUX achieves the top average Sharpe ratio with narrow bootstrap confidence intervals, maintains a flatter cost-performance slope than strong baselines, and attains superior risk-return efficiency for a given turnover budget. Pairwise scenario-level improvements are strictly positive and remain statistically significant after multiple-testing corrections. We provide formal guarantees on optimality under convex frictions, monotonic improvement under a KL trust region, long-run turnover bounds and induced inaction bands due to proportional costs, positive value advantage for regime-conditioned policies, and robustness to cost misspecification. The methodology is implementable: costs are calibrated from standard liquidity proxies, scenario-level inference avoids pseudo-replication, and all figures and tables are reproducible from released artifacts.

📄 PDF Abstract BibTeX arXiv:2510.02986

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Simulating Tenant Responses to Energy Policy Interventions with Transaction-Cost-Aware LLM Age

2026-07-27 · Weijie Xia, Stefanie Horian, Hanyue Huang, Queena K. Qian 외 arxiv

Recent studies use Large language models (LLMs) to simulate human opinions and decisions by prompting models with demographic, attitudinal, or persona-based descriptions. Yet such simulations rarely model the practical, …

Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning

2025-05-19 · Hongjoon Ahn, Heewoong Choi, Jisu Han, Taesup Moon

Offline goal-conditioned reinforcement learning (GCRL) offers a practical learning paradigm where goal-reaching policies are trained from abundant unlabeled (reward-free) datasets without additional environment interacti…

Uncovering Vulnerability of Vision-Language-Action Models under Joint-Level Physical Faults

2026-06-09 · Minsoo Jo, Taeju Kwon, Junha Chun, Youngjoon Jeong 외 arxiv

Deploying Vision-Language-Action (VLA) models in real robotic systems requires robustness not only to semantic and perceptual variations, but also to embodiment-side faults that change how actions are physically realized…

RecoverFormer: End-to-End Contact-Aware Recovery for Humanoid Robots

2026-04-24 · Zihui Liu arxiv

Humanoid robots operating in unstructured environments must recover from unexpected disturbances-a capability that remains challenging for end-to-end control policies. We present RECOVERFORMER, a fully end-to-end humanoi…

Auction-Based Task Allocation with Energy-Conscientious Trajectory Optimization for AMR Fleets

2026-03-23 · Jiachen Li, Soovadeep Bakshi, Jian Chu, Shihao Li 외 arxiv

This paper presents a hierarchical two-stage framework for multi-robot task allocation and trajectory optimization in asymmetric task spaces: (1) a sequential auction allocates tasks using closed-form bid functions, and …

Collision Avoidance