paper-with-me

홈 › Papers

Explaining Deep Neural Networks by Leveraging Intrinsic Methods

2024-07-17 · Biagio La Rosa

Despite their impact on the society, deep neural networks are often regarded as black-box models due to their intricate structures and the absence of explanations for their decisions. This opacity poses a significant challenge to AI systems wider adoption and trustworthiness. This thesis addresses this issue by contributing to the field of eXplainable AI, focusing on enhancing the interpretability of deep neural networks. The core contributions lie in introducing novel techniques aimed at making these networks more interpretable by leveraging an analysis of their inner workings. Specifically, the contributions are threefold. Firstly, the thesis introduces designs for self-explanatory deep neural networks, such as the integration of external memory for interpretability purposes and the usage of prototype and constraint-based layers across several domains. Secondly, this research delves into novel investigations on neurons within trained deep neural networks, shedding light on overlooked phenomena related to their activation values. Lastly, the thesis conducts an analysis of the application of explanatory techniques in the field of visual analytics, exploring the maturity of their adoption and the potential of these systems to convey explanations to users effectively.

📄 PDF Abstract BibTeX arXiv:2407.12243

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Intrinsic Barriers to Explaining Deep Foundation Models

2025-04-21 · Zhen Tan, Huan Liu

Deep Foundation Models (DFMs) offer unprecedented capabilities but their increasing complexity presents profound challenges to understanding their internal workings-a critical need for ensuring trust, safety, and account…

Leveraging Interpretable Tsetlin Machine for PDF Malware Detection

2026-07-10 · Rahul Jaiswal arxiv

In the digital era, Portable Document Format (PDF) is one of the most widely used file formats for storing and exchanging digital documents due to its platform independence and rich functionality. However, these same cap…

Computational EfficiencyMalware Detection

Self-Explaining Reinforcement Learning for Mobile Network Resource Allocation

2025-09-18 · Konrad Nowosadko, Franco Ruggeri, Ahmad Terra arxiv

Deep reinforcement learning (DRL) methods, though powerful, often lack transparency, which limits their adoption in critical domains. We apply Self-Explaining Neural Networks (SENNs) to RL by parametrizing the policy of …

Reinforcement Learning

Efficient Reinforcement Learning for Large Language Models with Intrinsic Exploration

2025-11-02 · Yan Sun, Jia Guo, Stanley Kok, Zihao Wang 외 arxiv

Reinforcement learning with verifiable rewards (RLVR) has improved the reasoning ability of large language models, yet training remains costly because many rollouts contribute little to optimization, considering the amou…

Reinforcement LearningMathematical Reasoning

Representation Gap: Explaining the Unreasonable Effectiveness of Neural Networks from a Geometric Perspective

2026-05-20 · David Perera, Victor Moura, Lais Isabelle Alves dos Santos, Michel F. C. Haddad 외 arxiv

Characterizing precisely the asymptotic generalization error of neural networks using parameters that can be estimated efficiently is a crucial problem in machine learning, which relies heavily on heuristics and practiti…