Representation Signatures and Risk-Feedback Alignment in LLM Trading Agents
We study behavioral alignment and representation dynamics of large language model (LLM) agents in financial decision environments. TradeArena, an auditable trading-agent testbed with risk reports, execution simulation, memory, and replayable trajectories, lets us analyze how rationales, positions, and interventions evolve under market stress. Code and data artifacts are available through the \href{https://github.com/weich97/TradeArena.git}{TradeArena repository}. We find pre-failure signatures: planning embeddings drift from normal centroids, fused plan-risk representations separate normal from pre-drawdown states, and local manifolds exhibit effective-rank contraction. Across 80 rolling failure anchors and eight LLM trajectories, this pattern persists across hash, LSA, Transformer, and white-box hidden-state probes. Stress tests with CoT-free target weights, lexical controls, OHLCV noise, and false audits show that rationale-level contraction can vanish without rationales, while intent-space and fused signatures remain informative. Structured risk feedback can act as an external alignment signal without fine-tuning, but not as a universal performance enhancer: true audit feedback improves calibration for some models, returns for others, and exposes cases where placebo or hidden feedback has higher short-horizon return but weaker alignment diagnostics. A 51-stock intraday experiment reveals a correlation blind spot: LLM rationales justify exposure to coupled assets that the risk layer clips. Finally, a financial-audit task suite shifts comparison from ``which model trades best'' to whether models can audit trajectories, respect execution boundaries, reproduce artifacts, and avoid claim overreach. These results support a research claim, not a profitability claim: auditable risk feedback and representation trajectories reveal when LLM financial reasoning is aligning, drifting, or failing.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Risk-aware Trading Portfolio Optimization
We investigate portfolio optimization in financial markets from a trading and risk management perspective. We term this task Risk-Aware Trading Portfolio Optimization (RATPO), formulate the corresponding optimization pro…
ManagementPortfolio OptimizationA model of adaptive, market behavior generating positive returns, volatility and system risk
We describe a simple model for speculative trading based on adaptive behavior of economic agents.The adaptive behavior is expressed through a feedback mechanism for changing agents' stock-to-bond ratios, depending on the…
Conditional Estimates of Diffusion Processes for Evaluating the Positive Feedback Trading
Positive feedback trading, which buys when prices rise and sells when prices fall, has long been criticized for being destabilizing as it moves prices away from the fundamentals. Motivated by the relationship between pos…
QuantAgents: Towards Multi-agent Financial System via Simulated Trading
In this paper, our objective is to develop a multi-agent financial system that incorporates simulated trading, a technique extensively utilized by financial professionals. While current LLM-based agent models demonstrate…
FinRLlama: A Solution to LLM-Engineered Signals Challenge at FinRL Contest 2024
In response to Task II of the FinRL Challenge at ACM ICAIF 2024, this study proposes a novel prompt framework for fine-tuning large language models (LLM) with Reinforcement Learning from Market Feedback (RLMF). Our frame…
Sentiment Analysis