paper-with-me

Papers

Delta-Audit: Explaining What Changes When Models Change

2025-08-27 · Arshia Hemmat, Afsaneh Fatemi arxiv

Model updates (new hyperparameters, kernels, depths, solvers, or data) change performance, but the \emph{reason} often remains opaque. We introduce \textbf{Delta-Attribution} (\mbox{$Δ$-Attribution}), a model-agnostic framework that explains \emph{what changed} between versions $A$ and $B$ by differencing per-feature attributions: $Δφ(x)=φ_B(x)-φ_A(x)$. We evaluate $Δφ$ with a \emph{$Δ$-Attribution Quality Suite} covering magnitude/sparsity (L1, Top-$k$, entropy), agreement/shift (rank-overlap@10, Jensen--Shannon divergence), behavioural alignment (Delta Conservation Error, DCE; Behaviour--Attribution Coupling, BAC; CO$Δ$F), and robustness (noise, baseline sensitivity, grouped occlusion). Instantiated via fast occlusion/clamping in standardized space with a class-anchored margin and baseline averaging, we audit 45 settings: five classical families (Logistic Regression, SVC, Random Forests, Gradient Boosting, $k$NN), three datasets (Breast Cancer, Wine, Digits), and three A/B pairs per family. \textbf{Findings.} Inductive-bias changes yield large, behaviour-aligned deltas (e.g., SVC poly$\!\rightarrow$rbf on Breast Cancer: BAC$\approx$0.998, DCE$\approx$6.6; Random Forest feature-rule swap on Digits: BAC$\approx$0.997, DCE$\approx$7.5), while ``cosmetic'' tweaks (SVC \texttt{gamma=scale} vs.\ \texttt{auto}, $k$NN search) show rank-overlap@10$=1.0$ and DCE$\approx$0. The largest redistribution appears for deeper GB on Breast Cancer (JSD$\approx$0.357). $Δ$-Attribution offers a lightweight update audit that complements accuracy by distinguishing benign changes from behaviourally meaningful or risky reliance shifts.

📄 PDF Abstract BibTeX arXiv:2508.19589

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Delta-XAI: A Unified Framework for Explaining Prediction Changes in Online Time Series Monitoring

2025-11-28 · Changhun Kim, Yechan Mun, Hyeongwon Jang, Eunseo Lee 외 arxiv

Explaining online time series monitoring models is crucial across sensitive domains such as healthcare and finance, where temporal and contextual prediction dynamics underpin critical decisions. While recent XAI methods …

Minimising changes to audit when updating decision trees

2024-08-29 · Anj Simmons, Scott Barnett, Anupam Chaudhuri, Sankhya Singh 외

Interpretable models are important, but what happens when the model is updated on new training data? We propose an algorithm for updating a decision tree while minimising the number of changes to the tree that a human wo…

DeltaSHAP: Explaining Prediction Evolutions in Online Patient Monitoring with Shapley Values

2025-07-03 · Changhun Kim, Yechan Mun, Sangchul Hahn, Eunho Yang arxiv

This study proposes DeltaSHAP, a novel explainable artificial intelligence (XAI) algorithm specifically designed for online patient monitoring systems. In clinical environments, discovering the causes driving patient ris…

Computational Efficiency

Signed Compression Progress on a Sealed Audit is Goodhart-Resistant

2026-06-09 · Ayush Mittal, Dhruv Gupta arxiv

Compression progress is a long-standing proposal for intrinsic motivation: reward an agent when its world model becomes better at predicting or compressing experience. The folk claim is that this reward is "credible" bec…

Exploring EEG Indicators to Evaluate Listening Difficulties in Noisy Environments

2025-06-17 · Azuki Onaya, Hiroki Tanaka

Auditory processing difficulties involve challenges in understanding speech in noisy environments despite normal hearing. However, the neural mechanisms remain unclear, and standardized diagnostic criteria are lacking. T…

DiagnosticEEG