paper-with-me

홈 › Papers

Evaluating Large Language Models for Stance Detection on Financial Targets from SEC Filing Reports and Earnings Call Transcripts

2025-10-27 · Nikesh Gyawali, Doina Caragea, Alex Vasenkov, Cornelia Caragea arxiv

Financial narratives from U.S. Securities and Exchange Commission (SEC) filing reports and quarterly earnings call transcripts (ECTs) are very important for investors, auditors, and regulators. However, their length, financial jargon, and nuanced language make fine-grained analysis difficult. Prior sentiment analysis in the financial domain required a large, expensive labeled dataset, making the sentence-level stance towards specific financial targets challenging. In this work, we introduce a sentence-level corpus for stance detection focused on three core financial metrics: debt, earnings per share (EPS), and sales. The sentences were extracted from Form 10-K annual reports and ECTs, and labeled for stance (positive, negative, neutral) using the advanced ChatGPT-o3-pro model under rigorous human validation. Using this corpus, we conduct a systematic evaluation of modern large language models (LLMs) using zero-shot, few-shot, and Chain-of-Thought (CoT) prompting strategies. Our results show that few-shot with CoT prompting performs best compared to supervised baselines, and LLMs' performance varies across the SEC and ECT datasets. Our findings highlight the practical viability of leveraging LLMs for target-specific stance in the financial domain without requiring extensive labeled data.

📄 PDF Abstract BibTeX arXiv:2510.23464

Code (0)

등록된 구현이 없습니다.

Tasks

Sentiment AnalysisStance Detection

Similar Papers 제목 키워드 기반

Won: Establishing Best Practices for Korean Financial NLP

2025-03-23 · Guijin Son, Hyunwoo Ko, Haneral Jung, Chami Hwang

In this work, we present the first open leaderboard for evaluating Korean large language models focused on finance. Operated for about eight weeks, the leaderboard evaluated 1,119 submissions on a closed benchmark coveri…

Stock Price Prediction

All That Glisters Is Not Gold: A Benchmark for Reference-Free Counterfactual Financial Misinformation Detection

2026-01-07 · Yuechen Jiang, Zhiwei Liu, Yupeng Cao, Yueru He 외 arxiv

We introduce RFC Bench, a benchmark for evaluating large language models on financial misinformation under realistic news. RFC Bench operates at the paragraph level and captures the contextual complexity of financial new…

EDINET-Bench: Evaluating LLMs on Complex Financial Tasks using Japanese Financial Statements

2025-06-10 · Issa Sugiura, Takashi Ishida, Taro Makino, Chieko Tazuke 외

Financial analysis presents complex challenges that could leverage large language model (LLM) capabilities. However, the scarcity of challenging financial datasets, particularly for Japanese financial data, impedes acade…

Binary ClassificationFinancial AnalysisFraud DetectionLarge Language Model

Artificial Intelligence-Enabled Accounting Information Systems and Fraud Detection in Nigeria's Financial Services Sector: The Moderating Role of Natural Language Processing

2026-06-04 · Timothy Oluwapelumi Adeyemi, Abigail Omotola Ojogbede arxiv

The rapid digitalisation of financial systems has improved operational efficiency and financial inclusion while simultaneously increasing exposure to sophisticated forms of cyber-enabled fraud and electronic financial mi…

Fraud Detection

Same Claim, Different Judgment: Benchmarking Scenario-Induced Bias in Multilingual Financial Misinformation Detection

2026-01-08 · Zhiwei Liu, Yupen Cao, Yuechen Jiang, Mohsinul Kabir 외 arxiv

Large language models (LLMs) have been widely applied across various domains of finance. Since their training data are largely derived from human-authored corpora, LLMs may inherit a range of human biases. Behavioral bia…