paper-with-me

Papers

Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining

2026-06-24 · Juliana Li, Diya Sreedhar arxiv

Midway through an ordinary pretraining run, a small language model learns the pronoun-gender rule: cued with a girl's name ("Sue cried because"), it resolves the next pronoun to she, generalizing to held-out probes (0.94 by step 925). By step 3,500 the same model scores near zero on the same probes, although the rule's evidence is still in the training data. We call this within-run reversal natural ungrokking: the corpus decides, with no trace in the loss curve, which learned rules a model keeps. Which rules survive is predictable from one corpus statistic: how often the training stream shows the rule winning. Across un-intervened runs (two corpora, three budgets, three seeds), support frequency decides a rule's fate; the data-to-parameter ratio only modulates how deeply a doomed rule falls. The same emerge-then-collapse dynamics appear in public Pythia checkpoints, collapse depth ordered by model scale as predicted. The forgetting is a displacement: a competing surface pattern out-competes the rule, and the log-probability margin between them crosses zero within 100 training steps of the behavioral collapse. Control over this fate is asymmetric: the same edit that destroys a rule on demand cannot restore it. Flipping support to counter-evidence in place kills the rule with monotone dose-response in two unrelated rules; but injecting support back, even to 450 times the level that naturally sustains it, buys no recovery. Every confirmatory threshold and prediction was pre-registered before the data it governed was read.

📄 PDF Abstract BibTeX arXiv:2606.26050

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Online Selective Conformal Prediction with Asymmetric Rules: A Permutation Test Approach

2026-02-10 · Mingyi Zheng, Ying Jin arxiv

Selective conformal prediction aims to construct prediction sets with valid coverage for a test unit conditional on it being selected by a data-driven mechanism. While existing methods in the offline setting handle any s…

Drug Discovery

Interpretable reinforcement learning for heat pump control through asymmetric differentiable decision trees

2025-06-02 · Toon Van Puyvelde, Mehran Zareh, Chris Develder

In recent years, deep reinforcement learning (DRL) algorithms have gained traction in home energy management systems. However, their adoption by energy management companies remains limited due to the black-box nature of …

Decision MakingDeep Reinforcement Learningenergy managementManagement+2

RuleCNL: A Controlled Natural Language for Business Rule Specifications

2014-06-09 · Paul Brillant Feuto Njonko, Sylviane Cardey, Peter Greenfield, Walid El Abed

Business rules represent the primary means by which companies define their business, perform their actions in order to reach their objectives. Thus, they need to be expressed unambiguously to avoid inconsistencies betwee…

Making Sense of Conflicting (Defeasible) Rules in the Controlled Natural Language ACE: Design of a System with Support for Existential Quantification Using Skolemization

2019-05-01 · WS 2019 5 · Martin Diller, Adam Wyner, Hannes Strass

We present the design of a system for making sense of conflicting rules expressed in a fragment of the prominent controlled natural language ACE, yet extended with means of expressing defeasible rules in the form of norm…

Translation

Decentralized Time and Energy-Optimal Control of Connected and Automated Vehicles in a Roundabout

2021-04-13 · Kaiyuan Xu, Christos G. Cassandras, Wei Xiao

The paper considers the problem of controlling Connected and Automated Vehicles (CAVs) traveling through a three-entry roundabout so as to jointly minimize both the travel time and the energy consumption while providing …