paper-with-me

Papers

X-SHIELD: Regularization for eXplainable Artificial Intelligence

2024-04-03 · Iván Sevillano-García, Julián Luengo, Francisco Herrera

As artificial intelligence systems become integral across domains, the demand for explainability grows, the called eXplainable artificial intelligence (XAI). Existing efforts primarily focus on generating and evaluating explanations for black-box models while a critical gap in directly enhancing models remains through these evaluations. It is important to consider the potential of this explanation process to improve model quality with a feedback on training as well. XAI may be used to improve model performance while boosting its explainability. Under this view, this paper introduces Transformation - Selective Hidden Input Evaluation for Learning Dynamics (T-SHIELD), a regularization family designed to improve model quality by hiding features of input, forcing the model to generalize without those features. Within this family, we propose the XAI - SHIELD(X-SHIELD), a regularization for explainable artificial intelligence, which uses explanations to select specific features to hide. In contrast to conventional approaches, X-SHIELD regularization seamlessly integrates into the objective function enhancing model explainability while also improving performance. Experimental validation on benchmark datasets underscores X-SHIELD's effectiveness in improving performance and overall explainability. The improvement is validated through experiments comparing models with and without the X-SHIELD regularization, with further analysis exploring the rationale behind its design choices. This establishes X-SHIELD regularization as a promising pathway for developing reliable artificial intelligence regularization.

📄 PDF Abstract BibTeX arXiv:2404.02611

Code (0)

등록된 구현이 없습니다.

Tasks

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors

2024-02-26 · Zhexin Zhang, Yida Lu, Jingyuan Ma, Di Zhang 외

The safety of Large Language Models (LLMs) has gained increasing attention in recent years, but there still lacks a comprehensive approach for detecting safety issues within LLMs' responses in an aligned, customizable an…

Explanation in Artificial Intelligence: Insights from the Social Sciences

2017-06-22 · Tim Miller

There has been a recent resurgence in the area of explainable artificial intelligence as researchers and practitioners seek to make their algorithms more understandable. Much of this research is focused on explicitly exp…

Explainable artificial intelligencePhilosophy

Deep Learning, Natural Language Processing, and Explainable Artificial Intelligence in the Biomedical Domain

2022-02-25 · Milad Moradi, Matthias Samwald

In this article, we first give an introduction to artificial intelligence and its applications in biology and medicine in Section 1. Deep learning methods are then described in Section 2. We narrow down the focus of the …

Explainable artificial intelligence

Foundations of Explainable Knowledge-Enabled Systems

2020-03-17 · Shruthi Chari, Daniel M. Gruen, Oshani Seneviratne, Deborah L. McGuinness

Explainability has been an important goal since the early days of Artificial Intelligence. Several approaches for producing explanations have been developed. However, many of these approaches were tightly coupled with th…

Explainable artificial intelligence

FakeShield: Explainable Image Forgery Detection and Localization via Multi-modal Large Language Models

2024-10-03 · Zhipei Xu, Xuanyu Zhang, Runyi Li, Zecheng Tang 외

The rapid development of generative AI is a double-edged sword, which not only facilitates content creation but also makes image manipulation easier and more difficult to detect. Although current image forgery detection …

Face SwappingImage Forgery DetectionImage ManipulationTAG