paper-with-me

홈 › Papers

Gradient Descent Robustly Learns the Intrinsic Dimension of Data in Training Convolutional Neural Networks

2025-04-11 · Chenyang Zhang, Peifeng Gao, Difan Zou, Yuan Cao

Modern neural networks are usually highly over-parameterized. Behind the wide usage of over-parameterized networks is the belief that, if the data are simple, then the trained network will be automatically equivalent to a simple predictor. Following this intuition, many existing works have studied different notions of "ranks" of neural networks and their relation to the rank of data. In this work, we study the rank of convolutional neural networks (CNNs) trained by gradient descent, with a specific focus on the robustness of the rank to image background noises. Specifically, we point out that, when adding background noises to images, the rank of the CNN trained with gradient descent is affected far less compared with the rank of the data. We support our claim with a theoretical case study, where we consider a particular data model to characterize low-rank clean images with added background noises. We prove that CNNs trained by gradient descent can learn the intrinsic dimension of clean images, despite the presence of relatively large background noises. We also conduct experiments on synthetic and real datasets to further validate our claim.

📄 PDF Abstract BibTeX arXiv:2504.08628

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Projected Stein Variational Gradient Descent

2020-02-09 · NeurIPS 2020 12 · Peng Chen, Omar Ghattas

The curse of dimensionality is a longstanding challenge in Bayesian inference in high dimensions. In this work, we propose a projected Stein variational gradient descent (pSVGD) method to overcome this challenge by explo…

Bayesian Inference

Projected Wasserstein gradient descent for high-dimensional Bayesian inference

2021-02-12 · Yifei Wang, Peng Chen, Wuchen Li

We propose a projected Wasserstein gradient descent method (pWGD) for high-dimensional Bayesian inference problems. The underlying density function of a particle system of WGD is approximated by kernel density estimation…

Bayesian InferenceDensity EstimationVocal Bursts Intensity Prediction

Securing Distributed Gradient Descent in High Dimensional Statistical Learning

2018-04-26 · Lili Su, Jiaming Xu

We consider unreliable distributed learning systems wherein the training data is kept confidential by external workers, and the learner has to interact closely with those workers to train a model. In particular, we assum…

Vocal Bursts Intensity Prediction

Grassmann Stein Variational Gradient Descent

2022-02-07 · Xing Liu, Harrison Zhu, Jean-François Ton, George Wynne 외

Stein variational gradient descent (SVGD) is a deterministic particle inference algorithm that provides an efficient alternative to Markov chain Monte Carlo. However, SVGD has been found to suffer from variance underesti…

Dimensionality Reduction

Learning Mixtures of Gaussians Using the DDPM Objective

2023-07-03 · NeurIPS 2023 11

Recent works have shown that diffusion models can learn essentially any distribution provided one can perform score estimation. Yet it remains poorly understood under what settings score estimation is possible, let alone…

Denoising