paper-with-me

홈 › Papers

From Neurons to Neutrons: A Case Study in Interpretability

2024-05-27 · Ouail Kitouni, Niklas Nolte, Víctor Samuel Pérez-Díaz, Sokratis Trifinopoulos, Mike Williams

Mechanistic Interpretability (MI) promises a path toward fully understanding how neural networks make their predictions. Prior work demonstrates that even when trained to perform simple arithmetic, models can implement a variety of algorithms (sometimes concurrently) depending on initialization and hyperparameters. Does this mean neuron-level interpretability techniques have limited applicability? We argue that high-dimensional neural networks can learn low-dimensional representations of their training data that are useful beyond simply making good predictions. Such representations can be understood through the mechanistic interpretability lens and provide insights that are surprisingly faithful to human-derived domain knowledge. This indicates that such approaches to interpretability can be useful for deriving a new understanding of a problem from models trained to solve it. As a case study, we extract nuclear physics concepts by studying models trained to reproduce nuclear data.

📄 PDF Abstract BibTeX arXiv:2405.17425

Code (1)

samuelperezdi/nuclr-icml 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Irradiation Tests for Commercial Off-the Shelf Components with Atmospheric-like Neutrons and Heavy-Ions

2023-11-29 · Paolo Branchini, Andrea Fabbri, Sacha Cormenier, Marco Bernardini 외

This paper presents the results of the irradiation, performed with atmospheric-like neutrons and heavy-ions, of Commercial Off-the Shelf Components (COTS), which can be used in space missions. In such cases, it is crucia…

NeutronStream: A Dynamic GNN Training Framework with Sliding Window for Graph Streams

2023-12-05 · Chaoyi Chen, Dechao Gao, Yanfeng Zhang, Qiange Wang 외

Existing Graph Neural Network (GNN) training frameworks have been designed to help developers easily create performant GNN implementations. However, most existing GNN frameworks assume that the input graphs are static, b…

Graph Neural Network

A Case Study on Concept Induction for Neuron-Level Interpretability in CNN

2026-02-27 · Moumita Sen Sarma, Samatha Ereshi Akkamahadevi, Pascal Hitzler arxiv

Deep Neural Networks (DNNs) have advanced applications in domains such as healthcare, autonomous systems, and scene understanding, yet the internal semantics of their hidden neurons remain poorly understood. Prior work i…

Scene UnderstandingScene Recognition

Adjusting the nuclear reactor's neutron transport and diffusion theory for an alternative description and modelling of postage or supplies delivery processes

2023-07-15 · Nick P. Petropoulos

There seems to exist significant similarities between a reactor system and a supply chain from collection to delivery. In the reactor case, neutrons are continuously produced and absorbed in nuclear fuel. In a supply sys…

Unity

Improvement studies on neutron-gamma separation in HPGe detectors by using neural networks

2013-04-11 · Serkan Akkoyun, Tuncay Bayram, S. Okan Kara

The neutrons emitted in heavy-ion fusion-evaporation (HIFE) reactions together with the gamma-rays cause unwanted backgrounds in gamma-ray spectra. Especially in the nuclear reactions, where relativistic ion beams (RIBs)…