paper-with-me

Papers

Subdominant Dense Clusters Allow for Simple Learning and High Computational Performance in Neural Networks with Discrete Synapses

2015-09-18 · Carlo Baldassi, Alessandro Ingrosso, Carlo Lucibello, Luca Saglietti, Riccardo Zecchina

We show that discrete synaptic weights can be efficiently used for learning in large scale neural systems, and lead to unanticipated computational performance. We focus on the representative case of learning random patterns with binary synapses in single layer networks. The standard statistical analysis shows that this problem is exponentially dominated by isolated solutions that are extremely hard to find algorithmically. Here, we introduce a novel method that allows us to find analytical evidence for the existence of subdominant and extremely dense regions of solutions. Numerical experiments confirm these findings. We also show that the dense regions are surprisingly accessible by simple learning protocols, and that these synaptic configurations are robust to perturbations and generalize better than typical solutions. These outcomes extend to synapses with multiple states and to deeper neural architectures. The large deviation measure also suggests how to design novel algorithmic schemes for optimization based on local entropy maximization.

📄 PDF Abstract BibTeX arXiv:1509.05753

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Binary perceptron: efficient algorithms can find solutions in a rare well-connected cluster

2021-11-04 · Emmanuel Abbe, Shuangping Li, Allan Sly

It was recently shown that almost all solutions in the symmetric binary perceptron are isolated, even at low constraint densities, suggesting that finding typical solutions is hard. In contrast, some algorithms have been…

Learning by Steering the Neural Dynamics: A Statistical Mechanics Perspective

2025-10-13 · Mattia Scardecchia arxiv

Despite the striking successes of deep neural networks trained with gradient-based optimization, these methods differ fundamentally from their biological counterparts. This gap raises key questions about how nature achie…

Typical and atypical solutions in non-convex neural networks with discrete and continuous weights

2023-04-26 · Carlo Baldassi, Enrico M. Malatesta, Gabriele Perugini, Riccardo Zecchina

We study the binary and continuous negative-margin perceptrons as simple non-convex neural network models learning random rules and associations. We analyze the geometry of the landscape of solutions in both models and f…

Identifying subdominant collective effects in a large motorway network

2022-02-15 · Shanshan Wang, Michael Schreckenberg, Thomas Guhr

In a motorway network, correlations between parts or, more precisely, between the sections of (different) motorways, are of considerable interest. Knowledge of flows and velocities on individual motorways is not sufficie…

ClusterAttention: A training-free speedup of bidirectional attention

2026-08-27 · Kasper Nordenram, Amelie Dittmann arxiv

This paper introduces ClusterAttention, a general training-free speedup of bidirectional attention layers. Existing sparse attention methods either rely on structure in the input, such as order in language or spatial pro…

Video Generation