paper-with-me

Papers

Improving generalization by regularizing in $L^2$ function space

2018-01-01 · ICLR 2018 1 · Ari S. Benjamin, Konrad Kording

Learning rules for neural networks necessarily include some form of regularization. Most regularization techniques are conceptualized and implemented in the space of parameters. However, it is also possible to regularize in the space of functions. Here, we propose to measure networks in an $L^2$ Hilbert space, and test a learning rule that regularizes the distance a network can travel through $L^2$-space each update. This approach is inspired by the slow movement of gradient descent through parameter space as well as by the natural gradient, which can be derived from a regularization term upon functional change. The resulting learning rule, which we call Hilbert-constrained gradient descent (HCGD), is thus closely related to the natural gradient but regularizes a different and more calculable metric over the space of functions. Experiments show that the HCGD is efficient and leads to considerably better generalization.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Spectral Regularization Allows Data-frugal Learning over Combinatorial Spaces

2022-10-05 · Amirali Aghazadeh, Nived Rajaraman, Tony Tu, Kannan Ramchandran

Data-driven machine learning models are being increasingly employed in several important inference problems in biology, chemistry, and physics which require learning over combinatorial spaces. Recent empirical evidence (…

High-Order Tensor Regularization With Application to Attribute Ranking

2018-06-01 · CVPR 2018 6 · Kwang In Kim, Juhyun Park, James Tompkin

When learning functions on manifolds, we can improve performance by regularizing with respect to the intrinsic manifold geometry rather than the ambient space. However, when regularizing tensor learning, calculating the …

AttributeVocal Bursts Intensity Prediction

TRAM: Bridging Trust Regions and Sharpness Aware Minimization

2023-10-05 · Tom Sherborne, Naomi Saphra, Pradeep Dasigi, Hao Peng

Sharpness-aware minimization (SAM) reports improving domain generalization by reducing the loss surface curvature in the parameter space. However, generalization during fine-tuning is often more dependent on the transfer…

Cross-Lingual TransferDomain GeneralizationLanguage ModelingLanguage Modelling+1

DropKAN: Regularizing KANs by masking post-activations

2024-07-17 · Mohammed Ghaith Altarabichi

We propose DropKAN (Dropout Kolmogorov-Arnold Networks) a regularization method that prevents co-adaptation of activation function weights in Kolmogorov-Arnold Networks (KANs). DropKAN functions by embedding the drop mas…

Kolmogorov-Arnold Networks

Regularizing Solutions to the MEG Inverse Problem Using Space-Time Separable Covariance Functions

2016-04-17 · Arno Solin, Pasi Jylänki, Jaakko Kauramäki, Tom Heskes 외

In magnetoencephalography (MEG) the conventional approach to source reconstruction is to solve the underdetermined inverse problem independently over time and space. Here we present how the conventional approach can be e…