paper-with-me

Papers

Foresight Learning for SEC Risk Prediction

2026-01-27 · Benjamin Turtel, Paul Wilczewski, Danny Franklin, Kris Skotheim arxiv

Risk disclosures in SEC filings describe potential adverse events but rarely quantify their likelihood, limiting their usefulness for probabilistic analysis. A central obstacle is the absence of large-scale, risk-level supervision linking disclosed risks to realized outcomes. We introduce a fully automated data generation pipeline that converts qualitative SEC risk disclosures into temporally grounded supervision using only public data. For each filing, the pipeline generates firm-specific, time-bounded risk queries from the Risk Factors section and labels them by automatically resolving outcomes against subsequent disclosures. Using this dataset of risk queries and outcomes grounded in SEC filings, we train a compact large language model to estimate the probability that a disclosed risk will materialize within a specified horizon. Despite its modest size, the resulting model substantially improves over pretrained and heuristic baselines, and outperforms frontier general-purpose models, including GPT-5, on probabilistic accuracy and calibration. More broadly, this work demonstrates that Foresight Learning enables scalable and fully automated training of domain-specific expert models using only raw, chronological, in-domain text -- without proprietary data, external corpora, or manual annotation. The resulting models achieve frontier-level performance while remaining deployable on a single GPU. This result suggests a general pathway for learning calibrated, decision-relevant signals from naturally occurring enterprise documents. To support transparency and reproducibility, we open-source the evaluation dataset used in this study. Evaluation Data: https://huggingface.co/datasets/LightningRodLabs/sec_risk_questions_test_set Data Generation Platform: https://lightningrod.ai/ SDK: https://github.com/lightning-rod-labs/lightningrod-python-sdk

📄 PDF Abstract BibTeX arXiv:2601.19189

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

From Prediction to Foresight: The Role of AI in Designing Responsible Futures

2025-11-26 · Maria Perez-Ortiz arxiv

In an era marked by rapid technological advancements and complex global challenges, responsible foresight has emerged as an essential framework for policymakers aiming to navigate future uncertainties and shape the futur…

Agent-Supported Foresight for AI Systemic Risks: AI Agents for Breadth, Experts for Judgment

2026-02-09 · Leon Fröhling, Alessandro Giaconia, Edyta Paulina Bogucka, Daniele Quercia arxiv

AI impact assessments often stress near-term risks because human judgment degrades over longer horizons, exemplifying the Collingridge dilemma: foresight is most needed when knowledge is scarcest. To address long-term sy…

ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI

2026-02-15 · Haibo Tong, Feifei Zhao, Linghao Feng, Ruoyu Wu 외 arxiv

Rapidly evolving AI exhibits increasingly strong autonomy and goal-directed capabilities, accompanied by derivative systemic risks that are more unpredictable, difficult to control, and potentially irreversible. However,…

Considering Human Factors in Risk Maps for Robust and Foresighted Driver Warning

2023-06-06 · Tim Puphal, Ryohei Hirano, Malte Probst, Raphael Wenzel 외

Driver support systems that include human states in the support process is an active research field. Many recent approaches allow, for example, to sense the driver's drowsiness or awareness of the driving situation. Howe…

Large Language Models for Medical Forecasting -- Foresight 2

2024-12-14 · Zeljko Kraljevic, Joshua Au Yeung, Daniel Bean, James Teo 외

Foresight 2 (FS2) is a large language model fine-tuned on hospital data for modelling patient timelines (GitHub 'removed for anon'). It can understand patients' clinical notes and predict SNOMED codes for a wide range of…

Language ModellingLarge Language Model