paper-with-me

홈 › Papers

A Multiplicative Neural Network Architecture: Locality and Regularity of Approximation

2026-02-06 · Hee-Sun Choi, Beom-Seok Han arxiv

We introduce a multiplicative neural network architecture in which multiplicative interactions constitute the fundamental representation, rather than appearing as auxiliary components within an additive model. We establish a universal approximation theorem for this architecture and analyze its approximation properties in terms of locality and regularity in Bessel potential spaces. To complement the theoretical results, we conduct numerical experiments on representative targets exhibiting sharp transition layers or pointwise loss of higher-order regularity. The experiments focus on the spatial structure of approximation errors and on regularity-sensitive quantities, in particular, the convergence of Zygmund-type seminorms. The results show that the proposed multiplicative architecture yields residual error structures that are more tightly aligned with regions of reduced regularity and exhibit more stable convergence in regularity-sensitive metrics. These results demonstrate that adopting a multiplicative representation format has concrete implications for the localization and regularity behavior of neural network approximations, providing a direct connection between architectural design and analytical properties of the approximating functions.

📄 PDF Abstract BibTeX arXiv:2602.06374

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Simplex-FEM Networks (SiFEN): Learning A Triangulated Function Approximator

2025-11-06 · Chaymae Yahyati, Ismail Lamaakal, Khalid El Makkaoui, Ibrahim Ouahbi 외 arxiv

We introduce Simplex-FEM Networks (SiFEN), a learned piecewise-polynomial predictor that represents f: R^d -> R^k as a globally C^r finite-element field on a learned simplicial mesh in an optionally warped input space. E…

Expressivity of Bi-Lipschitz Normalizing Flows: A Score-Based Diffusion Perspective

2026-05-07 · Meira Iske, Carola-Bibiane Schönlieb arxiv

Many normalizing flow architectures impose regularity constraints, yet their distributional approximation properties are not fully characterized. We study the expressivity of bi-Lipschitz normalizing flows through the le…

Deep Q-Learning on Hölder Spaces

2026-06-15 · Qian Qi arxiv

We study the operator-theoretic core of Q-learning in continuous-time stochastic control with continuous states and actions. In value-based reinforcement learning, each Q-learning or DQN update is built from a Bellman op…

Reinforcement Learning

Convergence of Stochastic Gradient Langevin Dynamics in the Lazy Training Regime

2025-10-24 · Noah Oberweis, Semih Cayci arxiv

Continuous-time models provide important insights into the training dynamics of optimization algorithms in deep learning. In this work, we establish a non-asymptotic convergence analysis of stochastic gradient Langevin d…

A Simple and Efficient Smoothing Method for Faster Optimization and Local Exploration

2020-12-01 · NeurIPS 2020 12 · Kevin Scaman, Ludovic Dos Santos, Merwan Barlier, Igor Colin

This work proposes a novel smoothing method, called Bend, Mix and Release (BMR), that extends two well-known smooth approximations of the convex optimization literature: randomized smoothing and the Moreau envelope. The …