paper-with-me

Papers

A Momentum Accelerated Algorithm for ReLU-based Nonlinear Matrix Decomposition

2024-02-04 · Qingsong Wang, Chunfeng Cui, Deren Han

Recently, there has been a growing interest in the exploration of Nonlinear Matrix Decomposition (NMD) due to its close ties with neural networks. NMD aims to find a low-rank matrix from a sparse nonnegative matrix with a per-element nonlinear function. A typical choice is the Rectified Linear Unit (ReLU) activation function. To address over-fitting in the existing ReLU-based NMD model (ReLU-NMD), we propose a Tikhonov regularized ReLU-NMD model, referred to as ReLU-NMD-T. Subsequently, we introduce a momentum accelerated algorithm for handling the ReLU-NMD-T model. A distinctive feature, setting our work apart from most existing studies, is the incorporation of both positive and negative momentum parameters in our algorithm. Our numerical experiments on real-world datasets show the effectiveness of the proposed model and algorithm. Moreover, the code is available at https://github.com/nothing2wang/NMD-TM.

📄 PDF Abstract BibTeX arXiv:2402.02442

Code (1)

nothing2wang/nmd-tm 공식 구현

Similar Papers 제목 키워드 기반

Accelerated Algorithms for Nonlinear Matrix Decomposition with the ReLU function

2023-05-15 · Giovanni Seraghiti, Atharva Awari, Arnaud Vandaele, Margherita Porcelli 외

In this paper, we study the following nonlinear matrix decomposition (NMD) problem: given a sparse nonnegative matrix $X$, find a low-rank matrix $\Theta$ such that $X \approx f(\Theta)$, where $f$ is an element-wise non…

An Efficient Alternating Algorithm for ReLU-based Symmetric Matrix Decomposition

2025-03-21 · Qingsong Wang

Symmetric matrix decomposition is an active research area in machine learning. This paper focuses on exploiting the low-rank structure of non-negative and sparse symmetric matrices via the rectified linear unit (ReLU) ac…

A Modular Analysis of Provable Acceleration via Polyak's Momentum: Training a Wide ReLU Network and a Deep Linear Network

2020-10-04 · Jun-Kun Wang, Chi-Heng Lin, Jacob Abernethy

Incorporating a so-called "momentum" dynamic in gradient descent methods is widely used in neural net training as it has been broadly observed that, at least empirically, it often leads to significantly faster convergenc…

Provable Accelerated Convergence of Nesterov's Momentum for Deep ReLU Neural Networks

2023-06-13 · Fangshuo Liao, Anastasios Kyrillidis

Current state-of-the-art analyses on the convergence of gradient descent for training neural networks focus on characterizing properties of the loss landscape, such as the Polyak-Lojaciewicz (PL) condition and the restri…

Open-Ended Question Answering

Momentum-based Distributed Resource Scheduling Optimization Subject to Sector-Bound Nonlinearity and Latency

2025-03-08 · Mohammadreza Doostmohammadian, Zulfiya R. Gabidullina, Hamid R. Rabiee

This paper proposes an accelerated consensus-based distributed iterative algorithm for resource allocation and scheduling. The proposed gradient-tracking algorithm introduces an auxiliary variable to add momentum towards…

Scheduling