paper-with-me

홈 › Papers

Convex Basins in Single-Index Model Loss Landscapes: Applications to Robust Recovery under Strong Adversarial Corruption

2026-05-28 · Santanu Das, Sagnik Chatterjee, Jatin Batra arxiv

We study the problem of robustly learning Gaussian Single Index Models (SIMs) in the presence of heavy-tailed noise and a constant fraction of adversarially corrupted covariates and responses. Prior work on robust recovery has considered settings such as linear regression (Pensia et al., JASA 2024), strictly monotonic link functions (Awasthi et al., NeurIPS 2022), and phase retrieval (Buna and Rebeschini, AISTATS 2025). However, these techniques do not extend to generic asymmetric non-monotonic link functions such as \textsc{GeLU} and \textsc{Swish}, which arise naturally as scalar primitives in modern gated neural architectures. We close this gap by giving the first robust recovery algorithm with near-linear sample and time complexity for generic non-monotonic link functions, thereby establishing the first robust recovery guarantees for a broad family of nonlinear SIMs for which \textit{no guarantees were previously known}. Our central contribution is a new structural understanding of the Gaussian squared-loss landscape under adversarial contamination. Crucially, we prove that for a broad class of nonlinear non-monotonic SIMs, a dimension-independent, constant-radius convex basin exists around the ground truth and is efficiently reachable via robust spectral initialization even under adversarial contamination. Prior works fail to establish both guarantees simultaneously, thereby either breaking down under adversarial contamination or failing to handle generic non-monotonic link functions. Together, these structural insights yield a principled warm start for robust gradient descent that provably converges to a final estimation error of $O(σ\sqrtε)$ in $\tilde{O}(nd)$ time with $\tilde{O}(d)$ samples, where $ε$ is the contamination fraction.

📄 PDF Abstract BibTeX arXiv:2605.29497

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Entropic Confinement and Mode Connectivity in Overparameterized Neural Networks

2025-12-06 · Luca Di Carlo, Chase Goddard, David J. Schwab arxiv

Modern neural networks exhibit a striking property: basins of attraction in the loss landscape are often connected by low-loss paths, yet optimization dynamics generally remain confined to a single convex basin and rarel…

Phase transitions reveal hierarchical structure in deep neural networks

2025-12-05 · Ibrahim Talha Ersoy, Andrés Fernando Cardozo Licha, Karoline Wiesner arxiv

Training Deep Neural Networks relies on the model converging on a high-dimensional, non-convex loss landscape toward a good minimum. Yet, much of the phenomenology of training remains ill understood. We focus on three se…

How Good is a Single Basin?

2024-02-05 · Kai Lion, Lorenzo Noci, Thomas Hofmann, Gregor Bachmann

The multi-modal nature of neural loss landscapes is often considered to be the main driver behind the empirical success of deep ensembles. In this work, we probe this belief by constructing various "connected" ensembles …

Neural network optimization strategies and the topography of the loss landscape

2026-02-24 · Jianneng Yu, Alexandre V. Morozov arxiv

Neural networks are trained by optimizing multi-dimensional sets of fitting parameters on non-convex loss landscapes. Low-loss regions of the landscapes correspond to the parameter sets that perform well on the training …

Feature-metric Loss for Self-supervised Learning of Depth and Egomotion

2020-07-21 · ECCV 2020 8 · Chang Shu, Kun Yu, Zhixiang Duan, Kuiyuan Yang

Photometric loss is widely used for self-supervised depth and egomotion estimation. However, the loss landscapes induced by photometric differences are often problematic for optimization, caused by plateau landscapes for…

Depth EstimationMonocular Depth EstimationSelf-Supervised LearningVisual Odometry