paper-with-me

홈 › Papers

Predicting Neural Network Accuracy from Weights

2020-02-26 · Thomas Unterthiner, Daniel Keysers, Sylvain Gelly, Olivier Bousquet, Ilya Tolstikhin

We show experimentally that the accuracy of a trained neural network can be predicted surprisingly well by looking only at its weights, without evaluating it on input data. We motivate this task and introduce a formal setting for it. Even when using simple statistics of the weights, the predictors are able to rank neural networks by their performance with very high accuracy (R2 score more than 0.98). Furthermore, the predictors are able to rank networks trained on different, unobserved datasets and with different architectures. We release a collection of 120k convolutional neural networks trained on four different datasets to encourage further research in this area, with the goal of understanding network training and performance better.

📄 PDF Abstract BibTeX arXiv:2002.11448

Code (1)

mostafaelaraby/generalization-gap-features-tensorflow tf

Similar Papers 제목 키워드 기반

Predicting Parameters in Deep Learning

2013-06-03 · NeurIPS 2013 12 · Misha Denil, Babak Shakibi, Laurent Dinh, Marc'Aurelio Ranzato 외

We demonstrate that there is significant redundancy in the parameterization of several deep learning models. Given only a few weight values for each feature it is possible to accurately predict the remaining values. More…

Deep Learning

Machine Learning Calabi-Yau Hypersurfaces

2021-12-12 · David S. Berman, Yang-Hui He, Edward Hirst

We revisit the classic database of weighted-P4s which admit Calabi-Yau 3-fold hypersurfaces equipped with a diverse set of tools from the machine-learning toolbox. Unsupervised techniques identify an unanticipated almost…

BIG-bench Machine LearningClustering

Hyper-Representations: Self-Supervised Representation Learning on Neural Network Weights for Model Characteristic Prediction

2021-10-28 · NeurIPS 2021 12 · Konstantin Schürholt, Dimche Kostadinov, Damian Borth

Self-Supervised Learning (SSL) has been shown to learn useful and information-preserving representations. Neural Networks (NNs) are widely applied, yet their weight space is still not fully understood. Therefore, we prop…

Representation LearningSelf-Supervised Learning

FLARE: A Data-Efficient Surrogate for Predicting Displacement Fields in Directed Energy Deposition

2026-04-17 · Kittipong Thiamchaiboonthawee, Ghadi Nehme, Ram Mohan Telikicherla, Jiawei Tian 외 arxiv

Directed energy deposition (DED) produces complex thermo-mechanical responses that can lead to distortion and reduced dimensional accuracy of a manufactured part. Thermo-mechanical finite element simulations are widely u…

HyperLoRA for PDEs

2023-08-18 · Ritam Majumdar, Vishal Jadhav, Anirudh Deodhar, Shirish Karande 외

Physics-informed neural networks (PINNs) have been widely used to develop neural surrogates for solutions of Partial Differential Equations. A drawback of PINNs is that they have to be retrained with every change in init…

Meta-Learningregression