paper-with-me

홈 › Papers

Learning through Internalization

2026-06-18 · Nikolaos Tsilivis, Nirmit Joshi, Marko Medvedev, Julia Kempe, Nati Srebro arxiv

We study internalization processes, by which neural-network-based systems absorb an explicit computational procedure into their own weights, and how they facilitate learning. We investigate how transformers internalize the simulation of semiautomata by internalizing chain-of-thought (CoT) tokens, which classes of semiautomata are harder to internalize, and expose the flip side of internalization, that is, a progressive degradation of out-of-distribution performance. We then provide the first provable analysis of successful internalization: for the task of learning parities, we show that a simplified one-layer transformer provably first learns the target with explicit CoT supervision and then internalizes the autoregressive generation as CoT tokens are progressively removed, learning to directly compute the parity. This task is computationally hard to learn from data without CoT supervision. Finally, we discuss how learning through internalization relates to the \textit{Positive Distribution Shift} phenomenon recently introduced by~\citet{Med+26}.

📄 PDF Abstract BibTeX arXiv:2606.20937

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

2026-06-03 · Jingwen Chen, Wenkai Yang, Shengda Fan, Wenbo Nie 외 arxiv

Experience internalization converts contextual experience from past interactions into reusable parametric capability, offering a promising path toward continual learning in large language models (LLMs). While prior work …

Continual Learning

Value Internalization: Learning and Generalizing from Social Reward

2024-07-19 · Frieda Rong, Max Kleiman-Weiner

Social rewards shape human behavior. During development, a caregiver guides a learner's behavior towards culturally aligned goals and values. How do these behaviors persist and generalize when the caregiver is no longer …

CogFlow: Bridging Perception and Reasoning through Knowledge Internalization for Visual Mathematical Problem Solving

2026-01-05 · Shuhang Chen, Yunqiu Xu, Junjie Xie, Aojun Lu 외 arxiv

Despite significant progress, multimodal large language models continue to struggle with visual mathematical problem solving. Some recent works recognize that visual perception is a bottleneck in visual mathematical reas…

Information ExtractionMathematical Reasoning

SKILLC: Learning Autonomous Skill Internalization in LLM Agents via Contrastive Credit Assignment

2026-05-27 · Hongxiang Lin, Zhirui Kuai, Erpeng Xue, Lei Wang arxiv

Structured skill prompts improve exploration in long-horizon agentic reinforcement learning (RL). Skill-augmented RL methods retain external skills at inference, while skill-internalization RL methods withdraw them durin…

Reinforcement Learning

From Exposure to Internalization: Dual-Stream Calibration for In-context Clinical Reasoning

2026-04-07 · Chuang Zhao, Hongke Zhao, Xiaofang Zhou, Xiaomeng Li arxiv

Contextual clinical reasoning demands robust inference grounded in complex, heterogeneous clinical records. While state-of-the-art fine-tuning, in-context learning (ICL), and retrieval-augmented generation (RAG) enable k…