paper-with-me

홈 › Papers

Nerva: a Truly Sparse Implementation of Neural Networks

2024-07-24 · Wieger Wesselink, Bram Grooten, Qiao Xiao, Cassio de Campos, Mykola Pechenizkiy

We introduce Nerva, a fast neural network library under development in C++. It supports sparsity by using the sparse matrix operations of Intel's Math Kernel Library (MKL), which eliminates the need for binary masks. We show that Nerva significantly decreases training time and memory usage while reaching equivalent accuracy to PyTorch. We run static sparse experiments with an MLP on CIFAR-10. On high sparsity levels like $99\%$, the runtime is reduced by a factor of $4\times$ compared to a PyTorch model using masks. Similar to other popular frameworks such as PyTorch and Keras, Nerva offers a Python interface for users to work with.

📄 PDF Abstract BibTeX arXiv:2407.17437

Code (1)

wiegerw/nerva 공식 구현 jax

Tasks

Math

Methods 이 논문이 사용한 방법론

Library 설명 없음

Similar Papers 제목 키워드 기반

Product Kanerva Machines: Factorized Bayesian Memory

2020-02-06 · Adam Marblestone, Yan Wu, Greg Wayne

An ideal cognitively-inspired memory system would compress and organize incoming items. The Kanerva Machine (Wu et al, 2018) is a Bayesian model that naturally implements online memory compression. However, the organizat…

Clustering

MINERVA: Evaluating Complex Video Reasoning

2025-05-01 · Arsha Nagrani, Sachit Menon, Ahmet Iscen, Shyamal Buch 외

Multimodal LLMs are turning their focus to video benchmarks, however most video benchmarks only provide outcome supervision, with no intermediate or interpretable reasoning steps. This makes it challenging to assess if m…

BenchmarkingTemporal Localization

Prevalence and recoverability of syntactic parameters in sparse distributed memories

2015-10-21 · Jeong Joon Park, Ronnel Boettcher, Andrew Zhao, Alex Mun 외

We propose a new method, based on Sparse Distributed Memory (Kanerva Networks), for studying dependency relations between different syntactic parameters in the Principles and Parameters model of Syntax. We store data of …

Relation

MINERVA-Cultural: A Benchmark for Cultural and Multilingual Long Video Reasoning

2026-01-15 · Darshan Singh, Arsha Nagrani, Kawshik Manikantan, Harman Singh 외 arxiv

Recent advancements in video models have shown tremendous progress, particularly in long video understanding. However, current benchmarks predominantly feature western-centric data and English as the dominant language, i…

Truly Sparse Neural Networks at Scale

2021-02-02 · Selima Curci, Decebal Constantin Mocanu, Mykola Pechenizkiyi

Recently, sparse training methods have started to be established as a de facto approach for training and inference efficiency in artificial neural networks. Yet, this efficiency is just in theory. In practice, everyone u…