paper-with-me

홈 › Papers

Towards Robust Influence Functions with Flat Validation Minima

2025-05-25 · Xichen Ye, Yifan Wu, Weizhong Zhang, Cheng Jin, Yifan Chen

The Influence Function (IF) is a widely used technique for assessing the impact of individual training samples on model predictions. However, existing IF methods often fail to provide reliable influence estimates in deep neural networks, particularly when applied to noisy training data. This issue does not stem from inaccuracies in parameter change estimation, which has been the primary focus of prior research, but rather from deficiencies in loss change estimation, specifically due to the sharpness of validation risk. In this work, we establish a theoretical connection between influence estimation error, validation set risk, and its sharpness, underscoring the importance of flat validation minima for accurate influence estimation. Furthermore, we introduce a novel estimation form of Influence Function specifically designed for flat validation minima. Experimental results across various tasks validate the superiority of our approach.

📄 PDF Abstract BibTeX arXiv:2505.19097

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

How to escape sharp minima with random perturbations

2023-05-25 · Kwangjun Ahn, Ali Jadbabaie, Suvrit Sra

Modern machine learning applications have witnessed the remarkable success of optimization algorithms that are designed to find flat minima. Motivated by this design choice, we undertake a formal study that (i) formulate…

Deforming the Loss Surface

2020-07-24 · Liangming Chen, Long Jin, Xiujuan Du, Shuai Li 외

In deep learning, it is usually assumed that the shape of the loss surface is fixed. Differently, a novel concept of deformation operator is first proposed in this paper to deform the loss surface, thereby improving the …

Entropic gradient descent algorithms and wide flat minima

2020-06-14 · ICLR 2021 1 · Fabrizio Pittorino, Carlo Lucibello, Christoph Feinauer, Gabriele Perugini 외

The properties of flat minima in the empirical risk landscape of neural networks have been debated for some time. Increasing evidence suggests they possess better generalization capabilities with respect to sharp ones. F…

Unveiling the structure of wide flat minima in neural networks

2021-07-02 · Carlo Baldassi, Clarissa Lauditi, Enrico M. Malatesta, Gabriele Perugini 외

The success of deep learning has revealed the application potential of neural networks across the sciences and opened up fundamental theoretical problems. In particular, the fact that learning algorithms based on simple …

Zeroth-Order Optimization Finds Flat Minima

2025-06-05 · Liang Zhang, Bingcong Li, Kiran Koshy Thekumparampil, Sewoong Oh 외

Zeroth-order methods are extensively used in machine learning applications where gradients are infeasible or expensive to compute, such as black-box attacks, reinforcement learning, and language model fine-tuning. Existi…

Binary ClassificationLanguage ModelingLanguage Modelling