paper-with-me

홈 › Papers

Harmonizing Feature Attributions Across Deep Learning Architectures: Enhancing Interpretability and Consistency

2023-07-05 · Md Abdul Kadir, Gowtham Krishna Addluri, Daniel Sonntag

Ensuring the trustworthiness and interpretability of machine learning models is critical to their deployment in real-world applications. Feature attribution methods have gained significant attention, which provide local explanations of model predictions by attributing importance to individual input features. This study examines the generalization of feature attributions across various deep learning architectures, such as convolutional neural networks (CNNs) and vision transformers. We aim to assess the feasibility of utilizing a feature attribution method as a future detector and examine how these features can be harmonized across multiple models employing distinct architectures but trained on the same data distribution. By exploring this harmonization, we aim to develop a more coherent and optimistic understanding of feature attributions, enhancing the consistency of local explanations across diverse deep-learning models. Our findings highlight the potential for harmonized feature attribution methods to improve interpretability and foster trust in machine learning applications, regardless of the underlying architecture.

📄 PDF Abstract BibTeX arXiv:2307.02150

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Similar Papers 제목 키워드 기반

Provably Better Explanations with Optimized Aggregation of Feature Attributions

2024-06-07 · Thomas Decker, Ananta R. Bhattarai, Jindong Gu, Volker Tresp 외

Using feature attributions for post-hoc explanations is a common practice to understand and verify the predictions of opaque machine learning models. Despite the numerous techniques available, individual methods often pr…

Harmonizing Feature Maps: A Graph Convolutional Approach for Enhancing Adversarial Robustness

2024-06-17 · Kejia Zhang, Juanjuan Weng, Junwei Wu, Guoqing Yang 외

The vulnerability of Deep Neural Networks to adversarial perturbations presents significant security concerns, as the imperceptible perturbations can contaminate the feature space and lead to incorrect predictions. Recen…

Adversarial Robustness

Intrinsic Explainability of Multimodal Learning for Crop Yield Prediction

2025-08-09 · Hiba Najjar, Deepak Pathak, Marlon Nuske, Andreas Dengel arxiv

Multimodal learning enables various machine learning tasks to benefit from diverse data sources, effectively mimicking the interplay of different factors in real-world applications, particularly in agriculture. While the…

Crop Yield Prediction

Visual Explanations via Iterated Integrated Attributions

2023-10-28 · ICCV 2023 1 · Oren Barkan, Yehonatan Elisha, Yuval Asher, Amit Eshel 외

We introduce Iterated Integrated Attributions (IIA) - a generic method for explaining the predictions of vision models. IIA employs iterative integration across the input image, the internal representations generated by …

Class-Dependent Perturbation Effects in Evaluating Time Series Attributions

2025-02-24 · Gregor Baer, Isel Grau, Chao Zhang, Pieter Van Gorp

As machine learning models become increasingly prevalent in time series applications, Explainable Artificial Intelligence (XAI) methods are essential for understanding their predictions. Within XAI, feature attribution m…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Time SeriesTime Series Classification