paper-with-me

홈 › Papers

Beyond Sight: Probing Alignment Between Image Models and Blind V1

2024-03-05 · Jacob Granley, Galen Pogoncheff, Alfonso Rodil, Leili Soo, Lily Marie Turkstra, Lucas Gil Nadolskis, Arantxa Alfaro Saez, Cristina Soto Sanchez, Eduardo Fernandez Jover, Michael Beyeler

Neural activity in the visual cortex of blind humans persists in the absence of visual stimuli. However, little is known about the preservation of visual representation capacity in these cortical regions, which could have significant implications for neural interfaces such as visual prostheses. In this work, we present a series of analyses on the shared representations between evoked neural activity in the primary visual cortex (V1) of a blind human with an intracortical visual prosthesis, and latent visual representations computed in deep neural networks (DNNs). In the absence of natural visual input, we examine two alternative forms of inducing neural activity: electrical stimulation and mental imagery. We first quantitatively demonstrate that latent DNN activations are aligned with neural activity measured in blind V1. On average, DNNs with higher ImageNet accuracy or higher sighted primate neural predictivity are more predictive of blind V1 activity. We further probe blind V1 alignment in ResNet-50 and propose a proof-of-concept approach towards interpretability of blind V1 neurons. The results of these studies suggest the presence of some form of natural visual processing in blind V1 during electrically evoked visual perception and present unique directions in mechanistically understanding and interfacing with blind V1.

📄 PDF Abstract BibTeX arXiv:2403.12990

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Dynamic Reflections: Probing Video Representations with Text Alignment

2025-11-04 · Tyler Zhu, Tengda Han, Leonidas Guibas, Viorica Pătrăucean 외 arxiv

The alignment of representations from different modalities has recently been shown to provide insights on the structural similarities and downstream capabilities of different encoders across diverse data types. While sig…

Pixels to Principles: Probing Intuitive Physics Understanding in Multimodal Language Models

2025-07-22 · Mohamad Ballout, Serwan Jassim, Elia Bruni arxiv

This paper presents a systematic evaluation of state-of-the-art multimodal large language models (MLLMs) on intuitive physics tasks using the GRASP and IntPhys 2 datasets. We assess the open-source models InternVL 2.5, Q…

Probing the Robustness of Large Language Models Safety to Latent Perturbations

2025-06-19 · Tianle Gu, Kexin Huang, Zongqi Wang, Yixu Wang 외

Safety alignment is a key requirement for building reliable Artificial General Intelligence. Despite significant advances in safety alignment, we observe that minor latent shifts can still trigger unsafe responses in ali…

DiagnosticSafety Alignment

Evaluating Image Editing with LLMs: A Comprehensive Benchmark and Intermediate-Layer Probing Approach

2026-03-20 · Shiqi Gao, Zitong Xu, Kang Fu, Huiyu Duan 외 arxiv

Evaluating text-guided image editing (TIE) methods remains a challenging problem, as reliable assessment should simultaneously consider perceptual quality, alignment with textual instructions, and preservation of origina…

Image Editing

Behind the Scene: Revealing the Secrets of Pre-trained Vision-and-Language Models

2020-05-15 · ECCV 2020 8 · Jize Cao, Zhe Gan, Yu Cheng, Licheng Yu 외

Recent Transformer-based large-scale pre-trained models have revolutionized vision-and-language (V+L) research. Models such as ViLBERT, LXMERT and UNITER have significantly lifted state of the art across a wide range of …

coreference-resolutionCoreference Resolutioncross-modal alignment