paper-with-me

Papers

A visual encoding model based on deep neural networks and transfer learning

2019-02-23 · Chi Zhang, Kai Qiao, Linyuan Wang, Li Tong, Guoen Hu, Ruyuan Zhang, Bin Yan

Background: Building visual encoding models to accurately predict visual responses is a central challenge for current vision-based brain-machine interface techniques. To achieve high prediction accuracy on neural signals, visual encoding models should include precise visual features and appropriate prediction algorithms. Most existing visual encoding models employ hand-craft visual features (e.g., Gabor wavelets or semantic labels) or data-driven features (e.g., features extracted from deep neural networks (DNN)). They also assume a linear mapping between feature representation to brain activity. However, it remains unknown whether such linear mapping is sufficient for maximizing prediction accuracy. New Method: We construct a new visual encoding framework to predict cortical responses in a benchmark functional magnetic resonance imaging (fMRI) dataset. In this framework, we employ the transfer learning technique to incorporate a pre-trained DNN (i.e., AlexNet) and train a nonlinear mapping from visual features to brain activity. This nonlinear mapping replaces the conventional linear mapping and is supposed to improve prediction accuracy on brain activity. Results: The proposed framework can significantly predict responses of over 20% voxels in early visual areas (i.e., V1-lateral occipital region, LO) and achieve unprecedented prediction accuracy. Comparison with Existing Methods: Comparing to two conventional visual encoding models, we find that the proposed encoding model shows consistent higher prediction accuracy in all early visual areas, especially in relatively anterior visual areas (i.e., V4 and LO). Conclusions: Our work proposes a new framework to utilize pre-trained visual features and train non-linear mappings from visual features to brain activity.

📄 PDF Abstract BibTeX arXiv:1902.08793

Code (0)

등록된 구현이 없습니다.

Tasks

PredictionTransfer Learning

Similar Papers 제목 키워드 기반

VEIL: How Visual Encoding Hijacking Induces Bias In Vision Models

2026-07-06 · Suranjana Sooraj, Xuyang Chen, Madhumitha Venkatesan, Dongyu Liu arxiv

Rendering time series as chart images for CNN-based classification has become increasingly common in time-series classification (TSC). However, it remains unclear whether models learn underlying temporal patterns or rely…

Visual Feature Encoding for GNNs on Road Networks

2022-03-02 · Oliver Stromann, Alireza Razavi, Michael Felsberg

In this work, we present a novel approach to learning an encoding of visual features into graph neural networks with the application on road network data. We propose an architecture that combines state-of-the-art vision …

Classificationimage-classificationImage ClassificationTransfer Learning

DRB-GAN: A Dynamic ResBlock Generative Adversarial Network for Artistic Style Transfer

2021-08-17 · ICCV 2021 10 · Wenju Xu, Chengjiang Long, Ruisheng Wang, Guanghui Wang

The paper proposes a Dynamic ResBlock Generative Adversarial Network (DRB-GAN) for artistic style transfer. The style code is modeled as the shared parameters for Dynamic ResBlocks connecting both the style encoding netw…

DecoderGenerative Adversarial NetworkStyle Transfer

Text-to-Image Generation via Implicit Visual Guidance and Hypernetwork

2022-08-17 · Xin Yuan, Zhe Lin, Jason Kuen, Jianming Zhang 외

We develop an approach for text-to-image generation that embraces additional retrieval images, driven by a combination of implicit visual guidance loss and generative objectives. Unlike most existing text-to-image genera…

DiversityImage GenerationRetrievalText to Image Generation+1

FCNR: Fast Compressive Neural Representation of Visualization Images

2024-07-23 · Yunfei Lu, Pengfei Gu, Chaoli Wang

We present FCNR, a fast compressive neural representation for tens of thousands of visualization images under varying viewpoints and timesteps. The existing NeRVI solution, albeit enjoying a high compression ratio, incur…

Image Compression