paper-with-me

홈 › Papers

Natural scene reconstruction from fMRI signals using generative latent diffusion

2023-03-09 · Furkan Ozcelik, Rufin VanRullen

In neural decoding research, one of the most intriguing topics is the reconstruction of perceived natural images based on fMRI signals. Previous studies have succeeded in re-creating different aspects of the visuals, such as low-level properties (shape, texture, layout) or high-level features (category of objects, descriptive semantics of scenes) but have typically failed to reconstruct these properties together for complex scene images. Generative AI has recently made a leap forward with latent diffusion models capable of generating high-complexity images. Here, we investigate how to take advantage of this innovative technology for brain decoding. We present a two-stage scene reconstruction framework called `Brain-Diffuser''. In the first stage, starting from fMRI signals, we reconstruct images that capture low-level properties and overall layout using a VDVAE (Very Deep Variational Autoencoder) model. In the second stage, we use the image-to-image framework of a latent diffusion model (Versatile Diffusion) conditioned on predicted multimodal (text and visual) features, to generate final reconstructed images. On the publicly available Natural Scenes Dataset benchmark, our method outperforms previous models both qualitatively and quantitatively. When applied to synthetic fMRI patterns generated from individual ROI (region-of-interest) masks, our trained model creates compelling `ROI-optimal'' scenes consistent with neuroscientific knowledge. Thus, the proposed methodology can have an impact on both applied (e.g. brain-computer interface) and fundamental neuroscience.

📄 PDF Abstract BibTeX arXiv:2303.05334

Code (1)

ozcelikfu/brain-diffuser 공식 구현 pytorch

Tasks

Brain Computer InterfaceBrain DecodingDescriptive

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.

Similar Papers 제목 키워드 기반

Reconstructing Natural Scenes from fMRI Patterns using BigBiGAN

2020-01-31 · Milad Mozafari, Leila Reddy, Rufin VanRullen

Decoding and reconstructing images from brain imaging data is a research area of high interest. Recent progress in deep generative neural networks has introduced new opportunities to tackle this problem. Here, we employ …

AttributeGenerative Adversarial NetworkImage Reconstruction

Mind Reader: Reconstructing complex images from brain activities

2022-09-30 · Sikun Lin, Thomas Sprague, Ambuj K Singh

Understanding how the brain encodes external stimuli and how these stimuli can be decoded from the measured brain activities are long-standing and challenging questions in neuroscience. In this paper, we focus on reconst…

MindShot: Multi-Shot Video Reconstruction from fMRI with LLM Decoding

2025-08-04 · Wenwen Zeng, Yonghuang Wu, Yifan Chen, Xuan Xie 외 arxiv

Reconstructing dynamic videos from fMRI is important for understanding visual cognition and enabling vivid brain-computer interfaces. However, current methods are critically limited to single-shot clips, failing to addre…

Video Reconstruction

Seeing Through the Brain: New Insights from Decoding Visual Stimuli with fMRI

2025-10-17 · Zheng Huang, Enpei Zhang, Weikang Qiu, Yinghao Cai 외 arxiv

Understanding how the brain encodes visual information is a central challenge in neuroscience and machine learning. A promising approach is to reconstruct visual stimuli, essentially images, from functional Magnetic Reso…

Image ReconstructionObject Detection

Hi-DREAM: Brain-Inspired Hierarchical Diffusion for fMRI-to-Image Reconstruction via ROI Encoder and VisuAl Mapping

2025-11-14 · Guowei Zhang, Yun Zhao, Kai Sun, Moein Khajehnejad 외 arxiv

Reconstructing natural images from fMRI requires bridging neural activity with both the structural and semantic representations used by modern generative models. Existing diffusion-based decoders often condition on a sin…

Image Reconstruction