paper-with-me

Papers

A Novel Unified Parametric Assumption for Nonconvex Optimization

2025-02-17 · Artem Riabinin, Ahmed Khaled, Peter Richtárik

Nonconvex optimization is central to modern machine learning, but the general framework of nonconvex optimization yields weak convergence guarantees that are too pessimistic compared to practice. On the other hand, while convexity enables efficient optimization, it is of limited applicability to many practical problems. To bridge this gap and better understand the practical success of optimization algorithms in nonconvex settings, we introduce a novel unified parametric assumption. Our assumption is general enough to encompass a broad class of nonconvex functions while also being specific enough to enable the derivation of a unified convergence theorem for gradient-based methods. Notably, by tuning the parameters of our assumption, we demonstrate its versatility in recovering several existing function classes as special cases and in identifying functions amenable to efficient optimization. We derive our convergence theorem for both deterministic and stochastic optimization, and conduct experiments to verify that our assumption can hold practically over optimization trajectories.

📄 PDF Abstract BibTeX arXiv:2502.12329

Code (1)

99991/cifar10-fast-simple 공식 구현 pytorch

Tasks

Stochastic Optimization

Similar Papers 제목 키워드 기반

Nonconvex Matrix Completion with Linearly Parameterized Factors

2020-03-29 · Ji Chen, Xiao-Dong Li, Zongming Ma

Techniques of matrix completion aim to impute a large portion of missing entries in a data matrix through a small portion of observed ones. In practice including collaborative filtering, prior information and special str…

Collaborative FilteringMatrix Completion

A Unified Analysis of Stochastic Gradient Methods for Nonconvex Federated Optimization

2020-06-12 · Zhize Li, Peter Richtárik

In this paper, we study the performance of a large family of SGD variants in the smooth nonconvex regime. To this end, we propose a generic and flexible assumption capable of accurate modeling of the second moment of the…

A unified convergence theory for adaptive first-order methods in the nonconvex case, including AdaNorm, full and diagonal AdaGrad and Muon

2026-04-19 · S. Gratton, Ph. L. Toint arxiv

A unified framework for first-order optimization algorithms fornonconvex unconstrained optimization is proposed that uses adaptivelypreconditioned gradients and includes popular methods such as full anddiagonal AdaGrad, …

Parametric Nonconvex Optimization via Convex Surrogates

2026-04-07 · Renzi Wang, Panagiotis Patrinos, Alberto Bemporad arxiv

This paper presents a novel learning-based approach to construct a surrogate problem that approximates a given parametric nonconvex optimization problem. The surrogate function is designed to be the minimum of a finite s…

Nonconvex Optimization Framework for Group-Sparse Feedback Linear-Quadratic Optimal Control: Penalty Approach

2025-07-24 · Lechen Feng, Xun Li, Yuan-Hua Ni arxiv

This paper develops a unified nonconvex optimization framework for the design of group-sparse feedback controllers in infinite-horizon linear-quadratic (LQ) problems. We address two prominent extensions of the classical …