paper-with-me

홈 › Papers

Spontaneous Kolmogorov-Arnold Geometry in Shallow MLPs

2025-09-15 · Michael H. Freedman, Michael Mulligan arxiv

The Kolmogorov-Arnold (KA) representation theorem constructs universal, but highly non-smooth inner functions (the first layer map) in a single (non-linear) hidden layer neural network. Such universal functions have a distinctive local geometry, a "texture," which can be characterized by the inner function's Jacobian $J({\mathbf{x}})$, as $\mathbf{x}$ varies over the data. It is natural to ask if this distinctive KA geometry emerges through conventional neural network optimization. We find that indeed KA geometry often is produced when training vanilla single hidden layer neural networks. We quantify KA geometry through the statistical properties of the exterior powers of $J(\mathbf{x})$: number of zero rows and various observables for the minor statistics of $J(\mathbf{x})$, which measure the scale and axis alignment of $J(\mathbf{x})$. This leads to a rough understanding for where KA geometry occurs in the space of function complexity and model hyperparameters. The motivation is first to understand how neural networks organically learn to prepare input data for later downstream processing and, second, to learn enough about the emergence of KA geometry to accelerate learning through a timely intervention in network hyperparameters. This research is the "flip side" of KA-Networks (KANs). We do not engineer KA into the neural network, but rather watch KA emerge in shallow MLPs.

📄 PDF Abstract BibTeX arXiv:2509.12326

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Scale-Agnostic Kolmogorov-Arnold Geometry in Neural Networks

2025-11-26 · Mathew Vanherreweghe, Michael H. Freedman, Keith M. Adams arxiv

Recent work by Freedman and Mulligan demonstrated that shallow multilayer perceptrons spontaneously develop Kolmogorov-Arnold geometric (KAG) structure during training on synthetic three-dimensional tasks. However, it re…

KAN: Kolmogorov-Arnold Networks

2024-04-30 · Ziming Liu, YiXuan Wang, Sachin Vaidya, Fabian Ruehle 외

Inspired by the Kolmogorov-Arnold representation theorem, we propose Kolmogorov-Arnold Networks (KANs) as promising alternatives to Multi-Layer Perceptrons (MLPs). While MLPs have fixed activation functions on nodes ("ne…

Kolmogorov-Arnold Networks

MLPs and KANs for data-driven learning in physical problems: A performance comparison

2025-04-15 · Raghav Pant, Sikan Li, Xingjian Li, Hassan Iqbal 외

There is increasing interest in solving partial differential equations (PDEs) by casting them as machine learning problems. Recently, there has been a spike in exploring Kolmogorov-Arnold Networks (KANs) as an alternativ…

Kolmogorov-Arnold Networks

Kolmogorov Arnold Informed neural network: A physics-informed deep learning framework for solving forward and inverse problems based on Kolmogorov Arnold Networks

2024-06-16 · Yizheng Wang, Jia Sun, Jinshuai Bai, Cosmin Anitescu 외

AI for partial differential equations (PDEs) has garnered significant attention, particularly with the emergence of Physics-informed neural networks (PINNs). The recent advent of Kolmogorov-Arnold Network (KAN) indicates…

FormKolmogorov-Arnold Networks

Kolmogorov-Arnold PointNet: Deep learning for prediction of fluid fields on irregular geometries

2024-08-06 · Ali Kashefi

Kolmogorov-Arnold Networks (KANs) have emerged as a promising alternative to traditional Multilayer Perceptrons (MLPs) in deep learning. KANs have already been integrated into various architectures, such as convolutional…

Kolmogorov-Arnold Networks