paper-with-me

Papers

A Special Case of Quadratic Extrapolation Under the Neural Tangent Kernel

2025-12-11 · Abiel Kim arxiv

It has been demonstrated both theoretically and empirically that the ReLU MLP tends to extrapolate linearly for an out-of-distribution evaluation point. The machine learning literature provides ample analysis with respect to the mechanisms to which linearity is induced. However, the analysis of extrapolation at the origin under the NTK regime remains a more unexplored special case. In particular, the infinite-dimensional feature map induced by the neural tangent kernel is not translationally invariant. This means that the study of an out-of-distribution evaluation point very far from the origin is not equivalent to the evaluation of a point very near the origin. And since the feature map is rotation invariant, these two special cases may represent the most canonically extreme bounds of ReLU NTK extrapolation. Ultimately, it is this loose recognition of the two special cases of extrapolation that motivate the discovery of quadratic extrapolation for an evaluation close to the origin.

📄 PDF Abstract BibTeX arXiv:2512.15749

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the infinite width limit of neural networks with a standard parameterization

2020-01-21 · Jascha Sohl-Dickstein, Roman Novak, Samuel S. Schoenholz, Jaehoon Lee

There are currently two parameterizations used to derive fixed kernels corresponding to infinite width neural networks, the NTK (Neural Tangent Kernel) parameterization and the naive standard parameterization. However, t…

A Unified Analysis of Multi-task Functional Linear Regression Models with Manifold Constraint and Composite Quadratic Penalty

2022-11-09 · Shiyuan He, Hanxuan Ye, Kejun He

This work studies the multi-task functional linear regression models where both the covariates and the unknown regression coefficients (called slope functions) are curves. For slope function estimation, we employ penaliz…

Multi-Task Learningregression

Extrapolation and Spectral Bias of Neural Nets with Hadamard Product: a Polynomial Net Study

2022-09-16 · Yongtao Wu, Zhenyu Zhu, Fanghui Liu, Grigorios G Chrysos 외

Neural tangent kernel (NTK) is a powerful tool to analyze training dynamics of neural networks and their generalization bounds. The study on NTK has been devoted to typical neural network architectures, but it is incompl…

Generalization BoundsPolynomial Neural Networks

Bridging the Gap between Constant Step Size Stochastic Gradient Descent and Markov Chains

2017-07-20 · Aymeric Dieuleveut, Alain Durmus, Francis Bach

We consider the minimization of an objective function given access to unbiased estimates of its gradient through stochastic gradient descent (SGD) with constant step-size. While the detailed analysis was only performed f…

Optimal lower bounds for logistic log-likelihoods

2024-10-14 · Niccolò Anceschi, Tommaso Rigon, Giacomo Zanella, Daniele Durante

The logit transform is arguably the most widely-employed link function beyond linear settings. This transformation routinely appears in regression models for binary data and provides, either explicitly or implicitly, a c…

Bayesian InferenceregressionVariational Inference