paper-with-me

홈 › Papers

Neural Scaling Laws of Deep ReLU and Deep Operator Network: A Theoretical Study

2024-10-01 · Hao liu, Zecheng Zhang, Wenjing Liao, Hayden Schaeffer

Neural scaling laws play a pivotal role in the performance of deep neural networks and have been observed in a wide range of tasks. However, a complete theoretical framework for understanding these scaling laws remains underdeveloped. In this paper, we explore the neural scaling laws for deep operator networks, which involve learning mappings between function spaces, with a focus on the Chen and Chen style architecture. These approaches, which include the popular Deep Operator Network (DeepONet), approximate the output functions using a linear combination of learnable basis functions and coefficients that depend on the input functions. We establish a theoretical framework to quantify the neural scaling laws by analyzing its approximation and generalization errors. We articulate the relationship between the approximation and generalization errors of deep operator networks and key factors such as network model size and training data size. Moreover, we address cases where input functions exhibit low-dimensional structures, allowing us to derive tighter error bounds. These results also hold for deep ReLU networks and other similar structures. Our results offer a partial explanation of the neural scaling laws in operator learning and provide a theoretical foundation for their applications.

📄 PDF Abstract BibTeX arXiv:2410.00357

Code (0)

등록된 구현이 없습니다.

Tasks

Operator learning

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Focus 설명 없음

Similar Papers 제목 키워드 기반

Scaling Law of Neural Koopman Operators

2026-02-23 · Abulikemu Abuduweili, Yuyang Pang, Feihan Li, Changliu Liu arxiv

Data-driven neural Koopman operator theory has emerged as a powerful tool for linearizing and controlling nonlinear robotic systems. However, the performance of these data-driven models fundamentally depends on the trade…

Singular Values for ReLU Layers

2018-12-06 · Sören Dittmer, Emily J. King, Peter Maass

Despite their prevalence in neural networks we still lack a thorough theoretical characterization of ReLU layers. This paper aims to further our understanding of ReLU layers by studying how the activation function ReLU i…

A Deep Learning Framework for Multi-Operator Learning: Architectures and Approximation Theory

2025-10-29 · Adrien Weihs, Jingmin Sun, Zecheng Zhang, Hayden Schaeffer arxiv

While many problems in machine learning focus on learning mappings between finite-dimensional spaces, scientific applications require approximating mappings between function spaces, i.e., operators. We study the problem …

Computational Efficiency

Least-Squares Neural Network (LSNN) Method For Scalar Nonlinear Hyperbolic Conservation Laws: Discrete Divergence Operator

2021-10-21 · Zhiqiang Cai, Jingshuang Chen, Min Liu

A least-squares neural network (LSNN) method was introduced for solving scalar linear and nonlinear hyperbolic conservation laws (HCLs) in [7, 6]. This method is based on an equivalent least-squares (LS) formulation and …

Numerical Integration

Scaling Laws for the Principled Design, Initialization, and Preconditioning of ReLU Networks

2020-01-01 · ICLR 2020 1 · Aaron Defazio, Leon Bottou

Abstract In this work, we describe a set of rules for the design and initialization of well-conditioned neural networks, guided by the goal of naturally balancing the diagonal blocks of the Hessian at the start of traini…