paper-with-me

홈 › Papers

No Spurious Local Minima: on the Optimization Landscapes of Wide and Deep Neural Networks

2020-09-28 · Johannes Lederer

Empirical studies suggest that wide neural networks are comparably easy to optimize, but mathematical support for this observation is scarce. In this paper, we analyze the optimization landscapes of deep learning with wide networks. We prove especially that constraint and unconstraint empirical-risk minimization over such networks has no spurious local minima. Hence, our theories substantiate the common belief that increasing network widths not only improves the expressiveness of deep-learning pipelines but also facilitates their optimizations.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Similar Papers 제목 키워드 기반

Noisy Low-rank Matrix Optimization: Geometry of Local Minima and Convergence Rate

2022-03-08 · Ziye Ma, Somayeh Sojoudi

This paper is concerned with low-rank matrix optimization, which has found a wide range of applications in machine learning. This problem in the special case of matrix sensing has been studied extensively through the not…

Escaping Local Minima Provably in Non-convex Matrix Sensing: A Deterministic Framework via Simulated Lifting

2026-02-05 · Tianqi Shen, Jinji Yang, Junze He, Kunhan Gao 외 arxiv

Low-rank matrix sensing is a fundamental yet challenging nonconvex problem whose optimization landscape typically contains numerous spurious local minima, making it difficult for gradient-based optimizers to converge to …

No Spurious Local Minima in Nonconvex Low Rank Problems: A Unified Geometric Analysis

2017-04-03 · ICML 2017 8 · Rong Ge, Chi Jin, Yi Zheng

In this paper we develop a new framework that captures the common landscape underlying the common non-convex low-rank matrix problems including matrix sensing, matrix completion and robust PCA. In particular, we show for…

Matrix Completion

Learn2Hop: Learned Optimization on Rough Landscapes

2021-07-20 · Amil Merchant, Luke Metz, Sam Schoenholz, Ekin Dogus Cubuk

Optimization of non-convex loss surfaces containing many local minima remains a critical problem in a variety of domains, including operations research, informatics, and material design. Yet, current techniques either re…

Efficient ExplorationMeta-Learning

Who is Afraid of Big Bad Minima? Analysis of gradient-flow in spiked matrix-tensor models

2019-12-01 · NeurIPS 2019 12 · Stefano Sarao Mannelli, Giulio Biroli, Chiara Cammarota, Florent Krzakala 외

Gradient-based algorithms are effective for many machine learning tasks, but despite ample recent effort and some progress, it often remains unclear why they work in practice in optimising high-dimensional non-convex fun…