paper-with-me

Papers

Comparing Post-Hoc Explainable AI Methods for Interpreting Black-Box EEG Models in Depression Detection

2026-05-27 · Antonia Šarčević, Nikolina Frid arxiv

Recent advances in deep learning have enabled increasingly accurate electroencephalography (EEG)-based classification of Major Depressive Disorder (MDD), but the decision-making processes of high-capacity models remain difficult to interpret. This study investigates multiple post-hoc explainability methods applied to an InceptionTime architecture trained for EEG-based MDD detection. The analysis includes Shapley-based, gradient-based, and perturbation-based attribution approaches: DeepSHAP, Integrated Gradients, GradCAM, Occlusion, and Permutation Feature Importance. Explainability analysis was performed within a subject-level stratified 5-fold cross-validation framework using global attribution aggregation across EEG segments and subjects. The evaluated methods revealed partially convergent attribution patterns, with recurring emphasis on frontal, temporal, and posterior EEG regions, particularly in the right hemisphere. Quantitative comparison demonstrated substantial agreement between gradient- and perturbation-based approaches, while DeepSHAP produced comparatively distinct attribution distributions. At the same time, variability between explainability methods highlighted the influence of methodological assumptions on the resulting explanations. Overall, the results suggest that different post-hoc explainability approaches capture partially overlapping relevance structures in EEG-based deep learning models for depression detection. Although the observed attribution patterns are broadly consistent with several previous EEG studies of MDD, the analysis should be interpreted as exploratory rather than evidence of definitive neurophysiological biomarkers or clinical applicability. The study highlights both the usefulness and limitations of post-hoc explainability for interpreting black-box EEG classifiers in psychiatric applications.

📄 PDF Abstract BibTeX arXiv:2605.28977

Code (0)

등록된 구현이 없습니다.

Tasks

Feature Importance

Similar Papers 제목 키워드 기반

Am I Building a White Box Agent or Interpreting a Black Box Agent?

2020-07-02 · Tom Bewley

The rule extraction literature contains the notion of a fidelity-accuracy dilemma: when building an interpretable model of a black box function, optimising for fidelity is likely to reduce performance on the underlying t…

Explainable artificial intelligence

Explainable AI model reveals disease-related mechanisms in single-cell RNA-seq data

2025-01-07 · Mohammad Usman, Olga Varea, Petia Radeva, Josep Canals 외

Neurodegenerative diseases (NDDs) are complex and lack effective treatment due to their poorly understood mechanism. The increasingly used data analysis from Single nucleus RNA Sequencing (snRNA-seq) allows to explore tr…

A Survey of Explainable AI in Deep Visual Modeling: Methods and Metrics

2023-01-31 · Naveed Akhtar

Deep visual models have widespread applications in high-stake domains. Hence, their black-box nature is currently attracting a large interest of the research community. We present the first survey in Explainable AI that …

From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms

2025-08-14 · Zhaokun Jiang, Ziyin Zhang arxiv

Recent advancements in machine learning have spurred growing interests in automated interpreting quality assessment. Nevertheless, existing research suffers from insufficient examination of language use quality, unsatisf…

Feature EngineeringData Augmentation

On the Value of Labeled Data and Symbolic Methods for Hidden Neuron Activation Analysis

2024-04-21 · Abhilekha Dalal, Rushrukh Rayan, Adrita Barua, Eugene Y. Vasserman 외

A major challenge in Explainable AI is in correctly interpreting activations of hidden neurons: accurate interpretations would help answer the question of what a deep learning system internally detects as relevant in the…

Explanation Generation