paper-with-me

Papers

Sharp Lower Bounds on the Approximation Rate of Shallow Neural Networks

2021-06-28 · Jonathan W. Siegel, Jinchao Xu

We consider the approximation rates of shallow neural networks with respect to the variation norm. Upper bounds on these rates have been established for sigmoidal and ReLU activation functions, but it has remained an important open problem whether these rates are sharp. In this article, we provide a solution to this problem by proving sharp lower bounds on the approximation rates for shallow neural networks, which are obtained by lower bounding the $L^2$-metric entropy of the convex hull of the neural network basis functions. In addition, our methods also give sharp lower bounds on the Kolmogorov $n$-widths of this convex hull, which show that the variation spaces corresponding to shallow neural networks cannot be efficiently approximated by linear methods. These lower bounds apply to both sigmoidal activation functions with bounded variation and to activation functions which are a power of the ReLU. Our results also quantify how much stronger the Barron spectral norm is than the variation norm and, combined with previous results, give the asymptotics of the $L^\infty$-metric entropy up to logarithmic factors in the case of the ReLU activation function.

📄 PDF Abstract BibTeX arXiv:2106.14997

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Sharp Bounds on the Approximation Rates, Metric Entropy, and $n$-widths of Shallow Neural Networks

2021-01-29 · Jonathan W. Siegel, Jinchao Xu

In this article, we study approximation properties of the variation spaces corresponding to shallow neural networks with a variety of activation functions. We introduce two main tools for estimating the metric entropy, a…

Sharp Lower Bounds for Linearized ReLU^k Approximation on the Sphere

2025-10-05 · Tong Mao, Jinchao Xu arxiv

We prove a saturation theorem for linearized shallow ReLU$^k$ neural networks on the unit sphere $\mathbb S^d$. For any antipodally quasi-uniform set of centers, if the target function has smoothness $r>\tfrac{d+2k+1}{2}…

Error bounds for approximations with deep ReLU networks

2016-10-03 · Dmitry Yarotsky

We study expressive power of shallow and deep neural networks with piece-wise linear activation functions. We establish new rigorous upper and lower bounds for the network complexity in the setting of approximations in S…

Shallow ReLU$^s$ Networks in $L^p$-Type and Sobolev Spaces: Approximation and Path-Norm Controlled Generalization

2026-05-18 · Weizhao Li, Fanghui Liu, Lei Shi arxiv

This paper studies approximation by shallow ReLU$^s$ networks, $σ_s(t)=\max\{0,t\}^s$, together with their generalization behavior under $\ell_1$ path-norm control. For the $L^p$-type integral spaces $\widetilde{\mathcal…

On best approximation by multivariate ridge functions with applications to generalized translation networks

2024-12-11 · Paul Geuchen, Palina Salanevich, Olov Schavemaker, Felix Voigtlaender

We prove sharp upper and lower bounds for the approximation of Sobolev functions by sums of multivariate ridge functions, i.e., functions of the form $\mathbb{R}^d \ni x \mapsto \sum_{k=1}^n h_k(A_k x) \in \mathbb{R}$ wi…