paper-with-me

홈 › Papers

Internal Representations of Vision Models Through the Lens of Frames on Data Manifolds

2022-11-19 · Henry Kvinge, Grayson Jorgenson, Davis Brown, Charles Godfrey, Tegan Emerson

While the last five years have seen considerable progress in understanding the internal representations of deep learning models, many questions remain. This is especially true when trying to understand the impact of model design choices, such as model architecture or training algorithm, on hidden representation geometry and dynamics. In this work we present a new approach to studying such representations inspired by the idea of a frame on the tangent bundle of a manifold. Our construction, which we call a neural frame, is formed by assembling a set of vectors representing specific types of perturbations of a data point, for example infinitesimal augmentations, noise perturbations, or perturbations produced by a generative model, and studying how these change as they pass through a network. Using neural frames, we make observations about the way that models process, layer-by-layer, specific modes of variation within a small neighborhood of a datapoint. Our results provide new perspectives on a number of phenomena, such as the manner in which training with augmentation produces model invariance or the proposed trade-off between adversarial training and model generalization.

📄 PDF Abstract BibTeX arXiv:2211.10558

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Decoding Vision Transformers: the Diffusion Steering Lens

2025-04-18 · Ryota Takatsuki, Sonia Joseph, Ippei Fujisawa, Ryota Kanai

Logit Lens is a widely adopted method for mechanistic interpretability of transformer-based language models, enabling the analysis of how internal representations evolve across layers by projecting them into the output v…

Analyzing Hierarchical Structure in Vision Models with Sparse Autoencoders

2025-05-21 · Matthew Lyle Olson, Musashi Hinck, Neale Ratzlaff, Changbai Li 외

The ImageNet hierarchy provides a structured taxonomy of object categories, offering a valuable lens through which to analyze the representations learned by deep vision models. In this work, we conduct a comprehensive an…

StructLens: A Structural Lens for Language Models via Maximum Spanning Trees

2026-02-10 · Haruki Sakajo, Frederikus Hudi, Yusuke Sakai, Hidetaka Kamigaito 외 arxiv

Language exhibits inherent structures, a property that explains both language acquisition and language change. Given this characteristic, we expect language models to manifest their own internal structures as well. While…

Language AcquisitionDependency Parsing

From Behavioral Performance to Internal Competence: Interpreting Vision-Language Models with VLM-Lens

2025-10-02 · Hala Sheta, Eric Huang, Shuyu Wu, Ilia Alenabi 외 arxiv

We introduce VLM-Lens, a toolkit designed to enable systematic benchmarking, analysis, and interpretation of vision-language models (VLMs) by supporting the extraction of intermediate outputs from any layer during the fo…

Connotation Frames of Power and Agency in Modern Films

2017-09-01 · EMNLP 2017 9 · Maarten Sap, Marcella Cindy Prasettio, Ari Holtzman, Hannah Rashkin 외

The framing of an action influences how we perceive its actor. We introduce connotation frames of power and agency, a pragmatic formalism organized using frame semantic representations, to model how different levels of p…