paper-with-me

홈 › Papers

Zero-Ablation Overstates Register Content Dependence in DINO Vision Transformers

2026-04-15 · Felipe Parodi, Jordan Matelsky, Melanie Segado arxiv

Zero-ablation -- replacing token activations with zero vectors -- is widely used to probe token function in vision transformers. Register zeroing in DINOv2+registers and DINOv3 produces large drops (up to $-36.6$\,pp classification, $-30.9$\,pp segmentation), suggesting registers are functionally indispensable. However, three replacement controls -- mean-substitution, noise-substitution, and cross-image register-shuffling -- preserve performance across classification, correspondence, and segmentation, remaining within ${\sim}1$\,pp of the unmodified baseline. Per-patch cosine similarity shows these replacements genuinely perturb internal representations, while zeroing causes disproportionately large perturbations, consistent with why it alone degrades tasks. We conclude that zero-ablation overstates dependence on exact register content. In the frozen-feature evaluations we test, performance depends on plausible register-like activations rather than on exact image-specific values. Registers nevertheless buffer dense features from \texttt{[CLS]} dependence and are associated with compressed patch geometry. These findings, including the replacement-control results, replicate at ViT-B scale.

📄 PDF Abstract BibTeX arXiv:2604.14433

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

At-Grok Is Not Converged:A Measurement-Validity Audit for Grokking Representation Metrics

2026-07-07 · Truong Xuan Khanh arxiv

On modular arithmetic, a network's embedding keeps compressing for tens of thousands of steps after it has already generalized. Reading effective rank at the grokking transition overstates the converged value by 3-5x on …

Progressive Disclosure for LLM-Maintained Wiki Knowledge Bases: a Preregistered Ablation

2026-07-06 · Theodore O. Cochran arxiv

LLM agents increasingly answer questions against knowledge bases they help maintain. A common intuition holds that progressive disclosure, a compact catalog plus a one-line summary per page so the agent loads only what i…

The Long Tail, Not the Front Page: Cold-Start Prediction of Crowd Highlight Salience

2026-06-10 · Kazuki Nakayashiki, Keisuke Watanabe arxiv

A social highlighter's most useful signal -- which passages a crowd of readers marks -- exists only for documents people have already read. Can the aggregate crowd salience of a document be predicted from its text before…

Mechanisms of Prompt-Induced Hallucination in Vision-Language Models

2026-01-08 · William Rudman, Michal Golovanevsky, Dana Arad, Yonatan Belinkov 외 arxiv

Large vision-language models (VLMs) are highly capable, yet often hallucinate by favoring textual prompts over visual evidence. We study this failure mode in a controlled object-counting setting, where the prompt oversta…

Machine Zygote: Causal Biparental Heredity Before Learning in a Germline--Soma Artificial Agent

2026-09-15 · Lyes Saad Saoud arxiv

Artificial ontogeny, developmental encodings, robot reproduction, and inherited controllers are established research directions, yet a narrower question remains: can a newborn artificial agent exhibit measurable biparent…