paper-with-me

홈 › Papers

On Explicit Super-Expressive Approximation for Neural Networks

2026-07-07 · Feng-Lei Fan, Ze-Yu Li, Chen-Yu Wang, Jian-Jun Wang arxiv

In this work, we investigate the fixed-architecture neural network approximation with explicit parameter bounds and elementary activations. While prior work demonstrated super-expressive approximation using fixed-size networks, they lack quantitative and non-asymptotic characterizations of parameter magnitude with respect to the approximation error. We resolve this issue by introducing the Chinese Remainder Theorem as a constructive encoding mechanism. For Lipschitz continuous functions on $[0,1]^D$, we construct a width-$\max\{D,4\}$, depth-$5$ network with explicit parameter-error trade-offs. For Hölder-smooth functions in $C^{r,γ}_A\left([0,1]^D\right)$, our fixed network of width $\max\{2D,\ D+5N+1\}$ and depth $r + 9$ achieves the parameter magnitude $\mathcal{P}$ bounded by $\log_2 \mathcal{P}=\mathcal{O}\bigl(\varepsilon^{-2D/(r+γ)}\log(1/\varepsilon)\bigr)$. This is the dual result compared to those in the parameter-bounded and architecture-unbounded paradigm.

📄 PDF Abstract BibTeX arXiv:2607.06781

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the Expressive Power of Transformers for Maxout Networks and Continuous Piecewise Linear Functions

2026-03-03 · Linyan Gu, Lihua Yang, Feng Zhou arxiv

Transformer networks have achieved remarkable empirical success across a wide range of applications, yet their theoretical expressive power remains insufficiently understood. In this paper, we study the expressive capabi…

Efficient Geometry-aware 3D Generative Adversarial Networks

2021-12-15 · CVPR 2022 1 · Eric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 외

Unsupervised generation of high-quality multi-view-consistent images and 3D shapes using only collections of single-view 2D photographs has been a long-standing challenge. Existing 3D GANs are either compute-intensive or…

3D geometryComputational EfficiencyNeural Rendering

Understanding the Expressive Power and Mechanisms of Transformer for Sequence Modeling

2024-02-01 · Mingze Wang, Weinan E

We conduct a systematic study of the approximation properties of Transformer for sequence modeling with long, sparse and complicated memory. We investigate the mechanisms through which different components of Transformer…

On the Optimal Expressive Power of ReLU DNNs and Its Application in Approximation with Kolmogorov Superposition Theorem

2023-08-10 · Juncai He

This paper is devoted to studying the optimal expressive power of ReLU deep neural networks (DNNs) and its application in approximation via the Kolmogorov Superposition Theorem. We first constructively prove that any con…

Expressive Power and Approximation Errors of Restricted Boltzmann Machines

2014-06-12 · NeurIPS 2011 12 · Guido Montufar, Johannes Rauh, Nihat Ay

We present explicit classes of probability distributions that can be learned by Restricted Boltzmann Machines (RBMs) depending on the number of units that they contain, and which are representative for the expressive pow…