paper-with-me

Papers

Neural Encoding and Decoding with Deep Learning for Dynamic Natural Vision

2017-11-14

Convolutional neural network (CNN) driven by image recognition has been shown to be able to explain cortical responses to static pictures at ventral-stream areas. Here, we further showed that such CNN could reliably predict and decode functional magnetic resonance imaging data from humans watching natural movies, despite its lack of any mechanism to account for temporal dynamics or feedback processing. Using separate data, encoding and decoding models were developed and evaluated for describing the bi-directional relationships be-tween the CNN and the brain. Through the encoding models, the CNN-predicted areas covered not only the ventral stream, but also the dorsal stream, albe-it to a lesser degree; single-voxel response was visualized as the specific pixel pattern that drove the response, revealing the distinct representation of individual cortical location; cortical activation was synthesized from natural images with high-throughput to map category representation, con-trast, and selectivity. Through the decoding models, fMRI signals were directly decoded to estimate the feature representations in both visual and semantic spaces, for direct visual reconstruction and seman-tic categorization, respectively. These results cor-roborate, generalize, and extend previous findings, and highlight the value of using deep learning, as an all-in-one model of the visual cortex, to understand and decode natural vision.

📄 PDF Abstract BibTeX arXiv:1608.03425

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FUSION: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding

2025-04-14 · Zheng Liu, Mengjie Liu, Jingzhou Chen, Jingwei Xu 외

We introduce FUSION, a family of multimodal large language models (MLLMs) with a fully vision-language alignment and integration paradigm. Unlike existing methods that primarily rely on late-stage modality interaction du…

Learning Encoding-Decoding Direction Pairs to Unveil Concepts of Influence in Deep Vision Networks

2025-09-28 · Alexandros Doumanoglou, Kurt Driessens, Dimitrios Zarpalas arxiv

Empirical evidence shows that deep vision networks often represent concepts as directions in latent space with concept information written along directional components in the vector representation of the input. However, …

FreqSelect: Frequency-Aware fMRI-to-Image Reconstruction

2025-05-18 · Junliang Ye, Lei Wang, Md Zakir Hossain

Reconstructing natural images from functional magnetic resonance imaging (fMRI) data remains a core challenge in natural decoding due to the mismatch between the richness of visual stimuli and the noisy, low resolution n…

Image Reconstruction

Accelerated Decoding of Centroid Positional Encoding for Instance Segmentation

2026-09-15 · Carmelo Scribano, Filippo Muzzini, Nedyalko Prisadnikov, Mohammad Mahdi 외 arxiv

Beyond model inference, the decoding stage, which converts raw network outputs into task-level representations, constitutes a significant portion of the execution cost. Despite its practical impact, prediction decoding h…

Instance Segmentation

Creative Language Encoding under Censorship

2018-08-01 · COLING 2018 8 · Heng Ji, Kevin Knight

People often create obfuscated language for online communication to avoid Internet censorship, share sensitive information, express strong sentiment or emotion, plan for secret actions, trade illegal products, or simply …

Position