paper-with-me

홈 › Papers

Dichotomy of Early and Late Phase Implicit Biases Can Provably Induce Grokking

2023-11-30 · Kaifeng Lyu, Jikai Jin, Zhiyuan Li, Simon S. Du, Jason D. Lee, Wei Hu

Recent work by Power et al. (2022) highlighted a surprising "grokking" phenomenon in learning arithmetic tasks: a neural net first "memorizes" the training set, resulting in perfect training accuracy but near-random test accuracy, and after training for sufficiently longer, it suddenly transitions to perfect test accuracy. This paper studies the grokking phenomenon in theoretical setups and shows that it can be induced by a dichotomy of early and late phase implicit biases. Specifically, when training homogeneous neural nets with large initialization and small weight decay on both classification and regression tasks, we prove that the training process gets trapped at a solution corresponding to a kernel predictor for a long time, and then a very sharp transition to min-norm/max-margin predictors occurs, leading to a dramatic change in test accuracy.

📄 PDF Abstract BibTeX arXiv:2311.18817

Code (1)

vfleaking/grokking-dichotomy 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Weight Decay 설명 없음

Similar Papers 제목 키워드 기반

Explicit vs. Implicit: Investigating Social Bias in Large Language Models through Self-Reflection

2025-01-04 · Yachao Zhao, Bo wang, Yan Wang

Large Language Models (LLMs) have been shown to exhibit various biases and stereotypes in their generated content. While extensive research has investigated bias in LLMs, prior work has predominantly focused on explicit …

Modeling Image Tone Dichotomy with the Power Function

2024-09-10 · Axel Martinez, Gustavo Olague, Emilio Hernandez

The primary purpose of this paper is to present the concept of dichotomy in image illumination modeling based on the power function. In particular, we review several mathematical properties of the power function to ident…

Image Enhancement

MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered

2025-04-10 · Imran Mirza, Cole Huang, Ishwara Vasista, Rohan Patil 외

Multi-agent systems, which consist of multiple AI models interacting within a shared environment, are increasingly used for persona-based interactions. However, if not carefully designed, these systems can reinforce impl…

Fairness

Actions Speak Louder than Words: Agent Decisions Reveal Implicit Biases in Language Models

2025-01-29 · YuXuan Li, Hirokazu Shirado, Sauvik Das

While advances in fairness and alignment have helped mitigate overt biases exhibited by large language models (LLMs) when explicitly prompted, we hypothesize that these models may still exhibit implicit biases when simul…

Decision MakingFairness

PEFTDebias : Capturing debiasing information using PEFTs

2023-12-01 · Sumit Agarwal, Aditya Srikanth Veerubhotla, Srijan Bansal

The increasing use of foundation models highlights the urgent need to address and eliminate implicit biases present in them that arise during pretraining. In this paper, we introduce PEFTDebias, a novel approach that emp…

parameter-efficient fine-tuning