paper-with-me

홈 › Papers

All Leaks Count, Some Count More: Interpretable Temporal Contamination Detection and Mitigation in LLM Backtesting

2026-02-19 · Zeyu Zhang, Ryan Chen, Bradly C. Stadie arxiv

Backtesting LLMs on resolved events assumes models reason only from pre-cutoff knowledge, yet pretrained models inevitably leak post-cutoff knowledge. We introduce a claim-level evaluation framework that decomposes prediction rationales into atomic claims and applies Shapley values to quantify each claim's decision impact, yielding \textbf{Shapley-DCLR} (\textbf{Shapley}-weighted \textbf{D}ecision-\textbf{C}ritical \textbf{L}eakage \textbf{R}ate) -- an interpretable metric measuring what fraction of decision-driving reasoning is contaminated. We further propose \textbf{TimeSPEC} (\textbf{Time}-\textbf{S}upervised \textbf{P}rediction with \textbf{E}xtracted \textbf{C}laims), an inference-time architecture that interleaves temporally-filtered retrieval with claim-level supervision, producing predictions grounded entirely in pre-cutoff evidence. Across three LLMs, the ablation experiments confirm retrieval and supervision are jointly necessary; and a three-task probe further illstrates that the performance cost of temporal enforcement scales with each task's reliance on post-cutoff information.

📄 PDF Abstract BibTeX arXiv:2602.17234

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Survey of Privacy-Preserving Model Explanations: Privacy Risks, Attacks, and Countermeasures

2024-03-31 · Thanh Tam Nguyen, Thanh Trung Huynh, Zhao Ren, Thanh Toan Nguyen 외

As the adoption of explainable AI (XAI) continues to expand, the urgency to address its privacy implications intensifies. Despite a growing corpus of research in AI privacy and explainability, there is little attention o…

Privacy PreservingSurvey

Information Leakages in the Green Bond Market

2025-04-04 · Darren Shannon, Jin Gong, Barry Sheehan

Public announcement dates are used in the green bond literature to measure equity market reactions to upcoming green bond issues. We find a sizeable number of green bond announcements were pre-dated by anonymous informat…

Generating Counterfactual Explanations Using Cardinality Constraints

2024-04-11 · Rubén Ruiz-Torrubiano

Providing explanations about how machine learning algorithms work and/or make particular predictions is one of the main tools that can be used to improve their trusworthiness, fairness and robustness. Among the most intu…

counterfactualFairness

Interpretable Credit Application Predictions With Counterfactual Explanations

2018-11-13 · Rory Mc Grath, Luca Costabello, Chan Le Van, Paul Sweeney 외

We predict credit applications with off-the-shelf, interchangeable black-box classifiers and we explain single predictions with counterfactual explanations. Counterfactual explanations expose the minimal changes required…

counterfactual

Self-Interpretable Time Series Prediction with Counterfactual Explanations

2023-06-09 · Jingquan Yan, Hao Wang

Interpretable time series prediction is crucial for safety-critical areas such as healthcare and autonomous driving. Most existing methods focus on interpreting predictions by assigning important scores to segments of ti…

Autonomous DrivingcounterfactualCounterfactual InferencePrediction+2