paper-with-me

Papers

Accelerating Deep Learning with Memcomputing

2018-01-01 · Haik Manukian, Fabio L. Traversa, Massimiliano Di Ventra

Restricted Boltzmann machines (RBMs) and their extensions, called 'deep-belief networks', are powerful neural networks that have found applications in the fields of machine learning and artificial intelligence. The standard way to training these models resorts to an iterative unsupervised procedure based on Gibbs sampling, called 'contrastive divergence' (CD), and additional supervised tuning via back-propagation. However, this procedure has been shown not to follow any gradient and can lead to suboptimal solutions. In this paper, we show an efficient alternative to CD by means of simulations of digital memcomputing machines (DMMs). We test our approach on pattern recognition using a modified version of the MNIST data set. DMMs sample effectively the vast phase space given by the model distribution of the RBM, and provide a very good approximation close to the optimum. This efficient search significantly reduces the number of pretraining iterations necessary to achieve a given level of accuracy, as well as a total performance gain over CD. In fact, the acceleration of pretraining achieved by simulating DMMs is comparable to, in number of iterations, the recently reported hardware application of the quantum annealing method on the same network and data set. Notably, however, DMMs perform far better than the reported quantum annealing results in terms of quality of the training. We also compare our method to advances in supervised training, like batch-normalization and rectifiers, that work to reduce the advantage of pretraining. We find that the memcomputing method still maintains a quality advantage ($>1\%$ in accuracy, and a $20\%$ reduction in error rate) over these approaches. Furthermore, our method is agnostic about the connectivity of the network. Therefore, it can be extended to train full Boltzmann machines, and even deep networks at once.

📄 PDF Abstract BibTeX arXiv:1801.00512

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Similar Papers 제목 키워드 기반

Memcomputing with membrane memcapacitive systems

2014-10-14 · Yuriy V. Pershin, Fabio L. Traversa, Massimiliano Di Ventra

We show theoretically that networks of membrane memcapacitive systems -- capacitors with memory made out of membrane materials -- can be used to perform a complete set of logic gates in a massively parallel way by simply…

Memcomputing and Swarm Intelligence

2014-08-28 · Y. V. Pershin, M. Di Ventra

We explore the relation between memcomputing, namely computing with and in memory, and swarm intelligence algorithms. In particular, we show that one can design memristive networks to solve short-path optimization proble…

Scheduling

Memcomputing NP-complete problems in polynomial time using polynomial resources and collective states

2014-11-18 · Fabio L. Traversa, Chiara Ramella, Fabrizio Bonani, Massimiliano Di Ventra

Memcomputing is a novel non-Turing paradigm of computation that uses interacting memory cells (memprocessors for short) to store and process information on the same physical platform. It was recently proved mathematicall…

A Memcomputing Pascaline

2015-03-16 · Y. V. Pershin, L. K. Castelano, F. Hartmann, V. Lopez-Richard 외

The original Pascaline was a mechanical calculator able to sum and subtract integers. It encodes information in the angles of mechanical wheels and through a set of gears, and aided by gravity, could perform the calculat…

Self-averaging of digital memcomputing machines

2023-01-20 · Daniel Primosch, Yuan-Hang Zhang, Massimiliano Di Ventra

Digital memcomputing machines (DMMs) are a new class of computing machines that employ non-quantum dynamical systems with memory to solve combinatorial optimization problems. Here, we show that the time to solution (TTS)…

Combinatorial Optimization