paper-with-me

홈 › Papers

Beyond the Black Box: Interpretability of LLMs in Finance

2025-05-14 · Hariom Tatsat, Ariye Shater

Large Language Models (LLMs) exhibit remarkable capabilities across a spectrum of tasks in financial services, including report generation, chatbots, sentiment analysis, regulatory compliance, investment advisory, financial knowledge retrieval, and summarization. However, their intrinsic complexity and lack of transparency pose significant challenges, especially in the highly regulated financial sector, where interpretability, fairness, and accountability are critical. As far as we are aware, this paper presents the first application in the finance domain of understanding and utilizing the inner workings of LLMs through mechanistic interpretability, addressing the pressing need for transparency and control in AI systems. Mechanistic interpretability is the most intuitive and transparent way to understand LLM behavior by reverse-engineering their internal workings. By dissecting the activations and circuits within these models, it provides insights into how specific features or components influence predictions - making it possible not only to observe but also to modify model behavior. In this paper, we explore the theoretical aspects of mechanistic interpretability and demonstrate its practical relevance through a range of financial use cases and experiments, including applications in trading strategies, sentiment analysis, bias, and hallucination detection. While not yet widely adopted, mechanistic interpretability is expected to become increasingly vital as adoption of LLMs increases. Advanced interpretability tools can ensure AI systems remain ethical, transparent, and aligned with evolving financial regulations. In this paper, we have put special emphasis on how these techniques can help unlock interpretability requirements for regulatory and compliance purposes - addressing both current needs and anticipating future expectations from financial regulators globally.

📄 PDF Abstract BibTeX arXiv:2505.24650

Code (0)

등록된 구현이 없습니다.

Tasks

FairnessHallucinationSentiment Analysis

Similar Papers 제목 키워드 기반

Facilitating Long Context Understanding via Supervised Chain-of-Thought Reasoning

2025-02-18 · Jingyang Lin, Andy Wong, Tian Xia, Shenghua He 외

Recent advances in Large Language Models (LLMs) have enabled them to process increasingly longer sequences, ranging from 2K to 2M tokens and even beyond. However, simply extending the input sequence length does not neces…

2kLong-Context Understanding

Where MLLMs Attend and What They Rely On: Explaining Autoregressive Token Generation

2025-09-26 · Ruoyu Chen, Xiaoqing Guo, Kangwei Liu, Siyuan Liang 외 arxiv

Multimodal large language models (MLLMs) have demonstrated remarkable capabilities in aligning visual inputs with natural language outputs. Yet, the extent to which generated tokens depend on visual modalities remains po…

Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs

2026-04-22 · Krishiv Agarwal, Ramneet Kaur, Colin Samplawski, Manoj Acharya 외 arxiv

Effective safety auditing of large language models (LLMs) demands tools that go beyond black-box probing and systematically uncover vulnerabilities rooted in model internals. We present a comprehensive, interpretability-…

Opening the Black Box of Financial AI with CLEAR-Trade: A CLass-Enhanced Attentive Response Approach for Explaining and Visualizing Deep Learning-Driven Stock Market Prediction

2017-09-05 · Devinder Kumar, Graham W. Taylor, Alexander Wong

Deep learning has been shown to outperform traditional machine learning algorithms across a wide range of problem domains. However, current deep learning algorithms have been criticized as uninterpretable "black-boxes" w…

Decision MakingDeep LearningPredictionStock Market Prediction

Deep learning interpretability for rough volatility

2024-11-28 · Bo Yuan, Damiano Brigo, Antoine Jacquier, Nicola Pede

Deep learning methods have become a widespread toolbox for pricing and calibration of financial models. While they often provide new directions and research results, their `black box' nature also results in a lack of int…

Deep Learning