paper-with-me

홈 › Papers

Logical Interpretations of Autoencoders

2019-11-26 · Anton Fuxjaeger, Vaishak Belle

The unification of low-level perception and high-level reasoning is a long-standing problem in artificial intelligence, which has the potential to not only bring the areas of logic and learning closer together but also demonstrate how abstract concepts might emerge from sensory data. Precisely because deep learning methods dominate perception-based learning, including vision, speech, and linguistic grammar, there is fast-growing literature on how to integrate symbolic reasoning and deep learning. Broadly, efforts seem to fall into three camps: those focused on defining a logic whose formulas capture deep learning, ones that integrate symbolic constraints in deep learning, and others that allow neural computations and symbolic reasoning to co-exist separately, to enjoy the strengths of both worlds. In this paper, we identify another dimension to this inquiry: what do the hidden layers really capture, and how can we reason about that logically? In particular, we consider autoencoders that are widely used for dimensionality reduction and inject a symbolic generative framework onto the feature layer. This allows us, among other things, to generate example images for a class to get a sense of what was learned. Moreover, the modular structure of the proposed model makes it possible to learn relations over multiple images at a time, as well as handle noisy labels. Our empirical evaluations show the promise of this inquiry.

📄 PDF Abstract BibTeX arXiv:1911.11629

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningDimensionality Reduction

Similar Papers 제목 키워드 기반

Interpreting and Steering Protein Language Models through Sparse Autoencoders

2025-02-13 · Edith Natalia Villegas Garcia, Alessio Ansuini

The rapid advancements in transformer-based language models have revolutionized natural language processing, yet understanding the internal mechanisms of these models remains a significant challenge. This paper explores …

An X-Ray Is Worth 15 Features: Sparse Autoencoders for Interpretable Radiology Report Generation

2024-10-04 · Ahmed Abdulaal, Hugo Fry, Nina Montaña-Brown, Ayodeji Ijishakin 외

Radiological services are experiencing unprecedented demand, leading to increased interest in automating radiology report generation. Existing Vision-Language Models (VLMs) suffer from hallucinations, lack interpretabili…

Language ModellingMultimodal Reasoning

Insights into a radiology-specialised multimodal large language model with sparse autoencoders

2025-07-17 · Kenza Bouzid, Shruthi Bannur, Felix Meissen, Daniel Coelho de Castro 외 arxiv

Interpretability can improve the safety, transparency and trust of AI models, which is especially important in healthcare applications where decisions often carry significant consequences. Mechanistic interpretability, p…

Remote sensing framework for geological mapping via stacked autoencoders and clustering

2024-04-02 · Sandeep Nagar, Ehsan Farahbakhsh, Joseph Awange, Rohitash Chandra

Supervised machine learning methods for geological mapping via remote sensing face limitations due to the scarcity of accurately labelled training data that can be addressed by unsupervised learning, such as dimensionali…

ClusteringDimensionality Reduction

A continuity of Markov blanket interpretations under the Free Energy Principle

2022-01-18 · Anil Seth, Tomasz Korbak, Alexander Tschantz

Bruineberg and colleagues helpfully distinguish between instrumental and ontological interpretations of Markov blankets, exposing the dangers of using the former to make claims about the latter. However, proposing a shar…