paper-with-me

홈 › Papers

Over-Alignment vs Over-Fitting: The Role of Feature Learning Strength in Generalization

2026-01-31 · Taesun Yeom, Taehyeok Ha, Jaeho Lee arxiv

Feature learning strength (FLS), i.e., the inverse of the effective output scaling of a model, plays a critical role in shaping the optimization dynamics of neural nets. While its impact has been extensively studied under the asymptotic regimes -- both in training time and FLS -- existing theory offers limited insight into how FLS affects generalization in practical settings, such as when training is stopped upon reaching a target training risk. In this work, we investigate the impact of FLS on generalization in deep networks under such practical conditions. Through empirical studies, we first uncover the emergence of an $\textit{optimal FLS}$ -- neither too small nor too large -- that yields substantial generalization gains. This finding runs counter to the prevailing intuition that stronger feature learning universally improves generalization. To explain this phenomenon, we develop a theoretical analysis of gradient flow dynamics in two-layer ReLU nets trained with logistic loss, where FLS is controlled via initialization scale. Our main theoretical result establishes the existence of an optimal FLS arising from a trade-off between two competing effects: An excessively large FLS induces an $\textit{over-alignment}$ phenomenon that degrades generalization, while an overly small FLS leads to $\textit{over-fitting}$.

📄 PDF Abstract BibTeX arXiv:2602.00827

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Visual Alignment Constraint for Continuous Sign Language Recognition

2021-04-06 · ICCV 2021 10 · Yuecong Min, Aiming Hao, Xiujuan Chai, Xilin Chen

Vision-based Continuous Sign Language Recognition (CSLR) aims to recognize unsegmented signs from image streams. Overfitting is one of the most critical problems in CSLR training, and previous works show that the iterati…

Sign Language Recognition

Rethinking Cross-Modal Fine-Tuning: Optimizing the Interaction Between Feature Alignment and Target Fitting

2026-01-26 · Trong Khiem Tran, Manh Cuong Dao, Phi Le Nguyen, Thao Nguyen Truong 외 arxiv

Adapting pre-trained models to unseen feature modalities has become increasingly important due to the growing need for cross-disciplinary knowledge integration. A key challenge here is how to align the representation of …

Adversarial Feature Distribution Alignment for Semi-Supervised Learning

2019-12-22 · Christoph Mayer, Matthieu Paul, Radu Timofte

Training deep neural networks with only a few labeled samples can lead to overfitting. This is problematic in semi-supervised learning where only a few labeled samples are available. In this paper, we show that a consequ…

RADE: Random Add-Drop Edge as a Regularizer

2026-05-30 · Danial Saber, Amirali Salehi-Abari arxiv

Graph Neural Networks (GNNs) suffer from overfitting and over-squashing of long-range information. Stochastic graph augmentations (e.g., edge deletion) regularize training against overfitting but can introduce train-infe…

A survey and classification of face alignment methods based on face models

2023-11-06 · Jagmohan Meher, Hector Allende-Cid, Torbjörn E. M. Nordling

A face model is a mathematical representation of the distinct features of a human face. Traditionally, face models were built using a set of fiducial points or landmarks, each point ideally located on a facial feature, i…

Face AlignmentFace Model