paper-with-me

홈 › Papers

Adaptivity of deep ReLU network for learning in Besov and mixed smooth Besov spaces: optimal rate and curse of dimensionality

2018-10-18 · ICLR 2019 5 · Taiji Suzuki

Deep learning has shown high performances in various types of tasks from visual recognition to natural language processing, which indicates superior flexibility and adaptivity of deep learning. To understand this phenomenon theoretically, we develop a new approximation and estimation error analysis of deep learning with the ReLU activation for functions in a Besov space and its variant with mixed smoothness. The Besov space is a considerably general function space including the Holder space and Sobolev space, and especially can capture spatial inhomogeneity of smoothness. Through the analysis in the Besov space, it is shown that deep learning can achieve the minimax optimal rate and outperform any non-adaptive (linear) estimator such as kernel ridge regression, which shows that deep learning has higher adaptivity to the spatial inhomogeneity of the target function than other estimators such as linear ones. In addition to this, it is shown that deep learning can avoid the curse of dimensionality if the target function is in a mixed smooth Besov space. We also show that the dependency of the convergence rate on the dimensionality is tight due to its minimax optimality. These results support high adaptivity of deep learning and its superior ability as a feature extractor.

📄 PDF Abstract BibTeX arXiv:1810.08033

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…

Similar Papers 제목 키워드 기반

Approximation and learning of anisotropic and mixed smooth functions by deep ReLU neural networks

2026-05-29 · Yunfei Yang, Jun Fan arxiv

This paper studies how efficiently deep ReLU neural networks can approximate and learn smooth functions. When the error is measured in $L^p([0,1]^d)$ norm and the approximator is a network with width $W$ and depth $L$, r…

Estimation error analysis of deep learning on the regression problem on the variable exponent Besov space

2020-09-23 · Kazuma Tsuji, Taiji Suzuki

Deep learning has achieved notable success in various fields, including image and speech recognition. One of the factors in the successful performance of deep learning is its high feature extraction ability. In this stud…

Deep Learningspeech-recognitionSpeech Recognition

Approximation of Smoothness Classes by Deep Rectifier Networks

2020-07-30 · Mazen Ali, Anthony Nouy

We consider approximation rates of sparsely connected deep rectified linear unit (ReLU) and rectified power unit (RePU) neural networks for functions in Besov spaces $B^\alpha_{q}(L^p)$ in arbitrary dimension $d$, on gen…

Sample Complexity of Offline Reinforcement Learning with Deep ReLU Networks

2021-03-11 · Thanh Nguyen-Tang, Sunil Gupta, Hung Tran-The, Svetha Venkatesh

Offline reinforcement learning (RL) leverages previously collected data for policy optimization without any further active exploration. Despite the recent interest in this problem, its theoretical results in neural netwo…

Offline RLreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Optimal Approximation Rates for Deep ReLU Neural Networks on Sobolev and Besov Spaces

2022-11-25 · Jonathan W. Siegel

Let $\Omega = [0,1]^d$ be the unit cube in $\mathbb{R}^d$. We study the problem of how efficiently, in terms of the number of parameters, deep neural networks with the ReLU activation function can approximate functions i…