paper-with-me

홈 › Papers

Collective Kernel EFT for Pre-activation ResNets

2026-04-17 · Hidetoshi Kawase, Toshihiro Ota arxiv

In finite-width deep neural networks, the empirical kernel $G$ evolves stochastically across layers. We develop a collective kernel effective field theory (EFT) for pre-activation ResNets based on a $G$-only closure hierarchy and diagnose its finite validity window. Exploiting the exact conditional Gaussianity of residual increments, we derive an exact stochastic recursion for $G$. Applying Gaussian approximations systematically yields a continuous-depth ODE system for the mean kernel $K_0$, the kernel covariance $V_4$, and the $1/n$ mean correction $K_{1,\mathrm{EFT}}$, which emerges diagrammatically as a one-loop tadpole correction. Numerically, $K_0$ remains accurate at all depths. However, the $V_4$ equation residual accumulates to an $O(1)$ error at finite time, primarily driven by approximation errors in the $G$-only transport term. Furthermore, $K_{1,\mathrm{EFT}}$ fails due to the breakdown of the source closure, which exhibits a systematic mismatch even at initialization. These findings highlight the limitations of $G$-only state-space reduction and suggest extending the state space to incorporate the sigma-kernel.

📄 PDF Abstract BibTeX arXiv:2604.15742

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Neural signature kernels as infinite-width-depth-limits of controlled ResNets

2023-03-30 · Nicola Muca Cirone, Maud Lemercier, Cristopher Salvi

Motivated by the paradigm of reservoir computing, we consider randomly initialized controlled ResNets defined as Euler-discretizations of neural controlled differential equations (Neural CDEs), a unified architecture whi…

Gaussian Processes

Kernel-Based Smoothness Analysis of Residual Networks

2020-09-21 · Tom Tirer, Joan Bruna, Raja Giryes

A major factor in the success of deep neural networks is the use of sophisticated architectures rather than the classical multilayer perceptron (MLP). Residual networks (ResNets) stand out among these powerful modern arc…

Deep Learning without Shortcuts: Shaping the Kernel with Tailored Rectifiers

2022-03-15 · ICLR 2022 4 · Guodong Zhang, Aleksandar Botev, James Martens

Training very deep neural networks is still an extremely challenging task. The common solution is to use shortcut connections and normalization layers, which are both crucial ingredients in the popular ResNet architectur…

Deep Learning

Why Do Deep Residual Networks Generalize Better than Deep Feedforward Networks? -- A Neural Tangent Kernel Perspective

2020-02-14 · Kaixuan Huang, Yuqing Wang, Molei Tao, Tuo Zhao

Deep residual networks (ResNets) have demonstrated better generalization performance than deep feedforward networks (FFNets). However, the theory behind such a phenomenon is still largely unknown. This paper studies this…

Why Do Deep Residual Networks Generalize Better than Deep Feedforward Networks? --- A Neural Tangent Kernel Perspective

2020-12-01 · NeurIPS 2020 12 · Kaixuan Huang, Yuqing Wang, Molei Tao, Tuo Zhao

Deep residual networks (ResNets) have demonstrated better generalization performance than deep feedforward networks (FFNets). However, the theory behind such a phenomenon is still largely unknown. This paper studies this…