paper-with-me

Papers

Teaching CORnet Human fMRI Representations for Enhanced Model-Brain Alignment

2024-07-15 · Zitong Lu, Yile Wang

Deep convolutional neural networks (DCNNs) have demonstrated excellent performance in object recognition and have been found to share some similarities with brain visual processing. However, the substantial gap between DCNNs and human visual perception still exists. Functional magnetic resonance imaging (fMRI) as a widely used technique in cognitive neuroscience can record neural activation in the human visual cortex during the process of visual perception. Can we teach DCNNs human fMRI signals to achieve a more brain-like model? To answer this question, this study proposed ReAlnet-fMRI, a model based on the SOTA vision model CORnet but optimized using human fMRI data through a multi-layer encoding-based alignment framework. This framework has been shown to effectively enable the model to learn human brain representations. The fMRI-optimized ReAlnet-fMRI exhibited higher similarity to the human brain than both CORnet and the control model in within-and across-subject as well as within- and across-modality model-brain (fMRI and EEG) alignment evaluations. Additionally, we conducted an in-depth analyses to investigate how the internal representations of ReAlnet-fMRI differ from CORnet in encoding various object dimensions. These findings provide the possibility of enhancing the brain-likeness of visual models by integrating human neural data, helping to bridge the gap between computer vision and visual neuroscience.

📄 PDF Abstract BibTeX arXiv:2407.10414

Code (0)

등록된 구현이 없습니다.

Tasks

EEGObject Recognition

Similar Papers 제목 키워드 기반

Reconstructing Retinal Visual Images from 3T fMRI Data Enhanced by Unsupervised Learning

2024-04-07 · Yujian Xiong, Wenhui Zhu, Zhong-Lin Lu, Yalin Wang

The reconstruction of human visual inputs from brain activity, particularly through functional Magnetic Resonance Imaging (fMRI), holds promising avenues for unraveling the mechanisms of the human visual system. Despite …

Generative Adversarial Network

Psychometry: An Omnifit Model for Image Reconstruction from Human Brain Activity

2024-03-29 · CVPR 2024 1 · Ruijie Quan, Wenguan Wang, Zhibo Tian, Fan Ma 외

Reconstructing the viewed images from human brain activity bridges human and computer vision through the Brain-Computer Interface. The inherent variability in brain function between individuals leads existing literature …

Brain Computer InterfaceImage ReconstructionMixture-of-ExpertsSpecificity

What Makes a Face Look like a Hat: Decoupling Low-level and High-level Visual Properties with Image Triplets

2024-09-03 · Maytus Piriyajitakonkij, Sirawaj Itthipuripat, Ian Ballard, Ioannis Pappas

In visual decision making, high-level features, such as object categories, have a strong influence on choice. However, the impact of low-level features on behavior is less understood partly due to the high correlation be…

Decision Making

Correlation Networks for Extreme Multi-label Text Classification

2022-08-23 · Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining 2022 8 · Guangxu Xun, Kishlay Jha, Jianhui Sun, Aidong Zhang

This paper develops the Correlation Networks (CorNet) architecture for the extreme multi-label text classification (XMTC) task, where the objective is to tag an input text sequence with the most relevant subset of labels…

ClassificationMulti Label Text ClassificationMulti-Label Text ClassificationTAG+2

NeuroCine: Decoding Vivid Video Sequences from Human Brain Activties

2024-02-02 · Jingyuan Sun, Mingxiao Li, Zijiao Chen, Marie-Francine Moens

In the pursuit to understand the intricacies of human brain's visual processing, reconstructing dynamic visual experiences from brain activities emerges as a challenging yet fascinating endeavor. While recent advancement…

Contrastive LearningSSIMVideo Generation