paper-with-me

홈 › Papers

Directional Convergence Analysis under Spherically Symmetric Distribution

2021-05-09 · Dachao Lin, Zhihua Zhang

We consider the fundamental problem of learning linear predictors (i.e., separable datasets with zero margin) using neural networks with gradient flow or gradient descent. Under the assumption of spherically symmetric data distribution, we show directional convergence guarantees with exact convergence rate for two-layer non-linear networks with only two hidden nodes, and (deep) linear networks. Moreover, our discovery is built on dynamic from the initialization without both initial loss and perfect classification constraint in contrast to previous works. We also point out and study the challenges in further strengthening and generalizing our results.

📄 PDF Abstract BibTeX arXiv:2105.03879

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Faster Directional Convergence of Linear Neural Networks under Spherically Symmetric Data

2021-12-01 · NeurIPS 2021 12 · Dachao Lin, Ruoyu Sun, Zhihua Zhang

In this paper, we study gradient methods for training deep linear neural networks with binary cross-entropy loss. In particular, we show global directional convergence guarantees from a polynomial rate to a linear rate f…

Analytical Expression for Spherically Symmetric Photoacoustic Sources: A Unified General Solution (Theoretical Analysis and Derivation)

2026-01-12 · Shuang Li, Yibing Wang, Yu Zhang, Changhui Li arxiv

Here we present a comprehensive derivation of the analytical expression for the spatiotemporal acoustic pressure generated by photoacoustic sources with spherically symmetric initial pressure distributions. Starting from…

Hierarchical Modeling of Local Image Features through L_p-Nested Symmetric Distributions

2009-12-01 · NeurIPS 2009 12 · Matthias Bethge, Eero P. Simoncelli, Fabian H. Sinz

We introduce a new family of distributions, called $L_p${\em -nested symmetric distributions}, whose densities access the data exclusively through a hierarchical cascade of $L_p$-norms. This class generalizes the family …

Greedy and Random Quasi-Newton Methods with Faster Explicit Superlinear Convergence

2021-12-01 · NeurIPS 2021 12 · Dachao Lin, Haishan Ye, Zhihua Zhang

In this paper, we follow Rodomanov and Nesterov’s work to study quasi-Newton methods. We focus on the common SR1 and BFGS quasi-Newton methods to establish better explicit (local) superlinear convergence rates. First, ba…

lp-Recovery of the Most Significant Subspace among Multiple Subspaces with Outliers

2010-12-18 · Gilad Lerman, Teng Zhang

We assume data sampled from a mixture of d-dimensional linear subspaces with spherically symmetric distributions within each subspace and an additional outlier component with spherically symmetric distribution within the…