paper-with-me

홈 › Papers

NSD-Imagery: A benchmark dataset for extending fMRI vision decoding methods to mental imagery

2025-06-07 · CVPR 2025 1 · Reese Kneeland, Paul S. Scotti, Ghislain St-Yves, Jesse Breedlove, Kendrick Kay, Thomas Naselaris

We release NSD-Imagery, a benchmark dataset of human fMRI activity paired with mental images, to complement the existing Natural Scenes Dataset (NSD), a large-scale dataset of fMRI activity paired with seen images that enabled unprecedented improvements in fMRI-to-image reconstruction efforts. Recent models trained on NSD have been evaluated only on seen image reconstruction. Using NSD-Imagery, it is possible to assess how well these models perform on mental image reconstruction. This is a challenging generalization requirement because mental images are encoded in human brain activity with relatively lower signal-to-noise and spatial resolution; however, generalization from seen to mental imagery is critical for real-world applications in medical domains and brain-computer interfaces, where the desired information is always internally generated. We provide benchmarks for a suite of recent NSD-trained open-source visual decoding models (MindEye1, MindEye2, Brain Diffuser, iCNN, Takagi et al.) on NSD-Imagery, and show that the performance of decoding methods on mental images is largely decoupled from performance on vision reconstruction. We further demonstrate that architectural choices significantly impact cross-decoding performance: models employing simple linear decoding architectures and multimodal feature decoding generalize better to mental imagery, while complex architectures tend to overfit visual training data. Our findings indicate that mental imagery datasets are critical for the development of practical applications, and establish NSD-Imagery as a useful resource for better aligning visual decoding methods with this goal.

📄 PDF Abstract BibTeX arXiv:2506.06898

Code (0)

등록된 구현이 없습니다.

Tasks

Image Reconstruction

Similar Papers 제목 키워드 기반

Seeing the imagined: a latent functional alignment in visual imagery decoding from fMRI data

2026-04-15 · Fabrizio Spera, Tommaso Boccato, Michal Olak, Sara Cammarota 외 arxiv

Recent progress in visual brain decoding from fMRI has been enabled by large-scale datasets such as the Natural Scenes Dataset (NSD) and powerful diffusion-based generative models. While current pipelines are primarily o…

Brain Decoding

Looking through the mind's eye via multimodal encoder-decoder networks

2024-09-27 · Arman Afrasiyabi, Erica Busch, Rahul Singh, Dhananjay Bhaskar 외

In this work, we explore the decoding of mental imagery from subjects using their fMRI measurements. In order to achieve this decoding, we first created a mapping between a subject's fMRI signals elicited by the videos t…

Decoder

Mind-to-Image: Projecting Visual Mental Imagination of the Brain from fMRI

2024-04-08 · Hugo Caselles-Dupré, Charles Mellerio, Paul Hérent, Alizée Lopez-Persem 외

The reconstruction of images observed by subjects from fMRI data collected during visual stimuli has made strong progress in the past decade, thanks to the availability of extensive fMRI datasets and advancements in gene…

Image Generation

MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery

2026-05-16 · Reese Kneeland, Cesar Kadir Torrico Villanueva, Jordyn Ojeda, Shuhb Khanna 외 arxiv

To be useful for downstream applications, vision decoding models that are trained to reconstruct seen images from human brain activity must be able to generalize to internally generated visual representations, i.e., ment…

Image Reconstruction

fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding

2025-11-24 · Yuxiang Wei, Yanteng Zhang, Xi Xiao, Chengxuan Qian 외 arxiv

Recent advances in multimodal large language models (LLMs) have enabled unified reasoning across images, audio, and video, but extending such capability to brain imaging remains largely unexplored. Bridging this gap is e…