paper-with-me

홈 › Papers

Tracking the Feature Dynamics in LLM Training: A Mechanistic Study

2024-12-23 · Yang Xu, Yi Wang, Hao Wang

Understanding training dynamics and feature evolution is crucial for the mechanistic interpretability of large language models (LLMs). Although sparse autoencoders (SAEs) have been used to identify features within LLMs, a clear picture of how these features evolve during training remains elusive. In this study, we: (1) introduce SAE-Track, a novel method to efficiently obtain a continual series of SAEs; (2) mechanistically investigate feature formation and develop a progress measure for it ; and (3) analyze and visualize feature drift during training. Our work provides new insights into the dynamics of features in LLMs, enhancing our understanding of training mechanisms and feature evolution.

📄 PDF Abstract BibTeX arXiv:2412.17626

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Revealing the Learning Dynamics of Long-Context Continual Pre-training

2026-04-03 · Yupu Liang, Shuang Chen, Guanwei Zhang, Shaolei Wang 외 arxiv

Existing studies on Long-Context Continual Pre-training (LCCP) mainly focus on small-scale models and limited data regimes (tens of billions of tokens). We argue that directly migrating these small-scale settings to indu…

Residual Koopman Model Predictive Control for Enhanced Vehicle Dynamics with Small On-Track Data Input

2025-07-24 · Yonghao Fu, Cheng Hu, Haokun Xiong, Zhanpeng Bao 외 arxiv

In vehicle trajectory tracking tasks, the simplest approach is the Pure Pursuit (PP) Control. However, this single-point preview tracking strategy fails to consider vehicle model constraints, compromising driving safety.…

Computational Efficiency

Cochlear Wave Propagation and Dynamics in the Human Base and Apex: Model-Based Estimates from Noninvasive Measurements

2024-04-10 · Samiya A Alkhairy

Cochlear wavenumber and impedance are mechanistic variables that encode information regarding how the cochlea works - specifically wave propagation and Organ of Corti dynamics. These mechanistic variables underlie intere…

RADAR: Mechanistic Pathways for Detecting Data Contamination in LLM Evaluation

2025-10-10 · Ashish Kattamuri, Harshwardhan Fartale, Arpita Vats, Rahul Raja 외 arxiv

Data contamination poses a significant challenge to reliable LLM evaluation, where models may achieve high performance by memorizing training data rather than demonstrating genuine reasoning capabilities. We introduce RA…

Accurate Open-Loop Control of a Soft Continuum Robot Through Visually Learned Latent Representations

2026-03-20 · Henrik Krauss, Johann Licher, Naoya Takeishi, Annika Raatz 외 arxiv

This work addresses open-loop control of a soft continuum robot (SCR) from video-learned latent dynamics. Visual Oscillator Networks (VONs) from previous work are used, that provide mechanistically interpretable 2D oscil…