paper-with-me

홈 › Papers

Human-Aligned Image Models Improve Visual Decoding from the Brain

2025-02-05 · Nona Rajabi, Antônio H. Ribeiro, Miguel Vasco, Farzaneh Taleb, Mårten Björkman, Danica Kragic

Decoding visual images from brain activity has significant potential for advancing brain-computer interaction and enhancing the understanding of human perception. Recent approaches align the representation spaces of images and brain activity to enable visual decoding. In this paper, we introduce the use of human-aligned image encoders to map brain signals to images. We hypothesize that these models more effectively capture perceptual attributes associated with the rapid visual stimuli presentations commonly used in visual brain data recording experiments. Our empirical results support this hypothesis, demonstrating that this simple modification improves image retrieval accuracy by up to 21% compared to state-of-the-art methods. Comprehensive experiments confirm consistent performance improvements across diverse EEG architectures, image encoders, alignment methods, participants, and brain imaging modalities.

📄 PDF Abstract BibTeX arXiv:2502.03081

Code (0)

등록된 구현이 없습니다.

Tasks

EEGImage RetrievalRetrieval

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Learning to Expand Images for Efficient Visual Autoregressive Modeling

2025-11-19 · Ruiqing Yang, Kaixin Zhang, Zheng Zhang, Shan You 외 arxiv

Autoregressive models have recently shown great promise in visual generation by leveraging discrete token sequences akin to language modeling. However, existing approaches often suffer from inefficiency, either due to to…

Image Generation

Brain-aligning of semantic vectors improves neural decoding of visual stimuli

2024-03-22 · Shirin Vafaei, Ryohei Fukuma, Takufumi Yanagisawa, Huixiang Yang 외

The development of algorithms to accurately decode of neural information is a long-standing effort in the field of neuroscience. Brain decoding is typically employed by training machine learning models to map neural data…

Brain DecodingRepresentation Learning

DuetSVG: Unified Multimodal SVG Generation with Internal Visual Guidance

2025-12-11 · Peiying Zhang, Nanxuan Zhao, Matthew Fisher, Yiran Xu 외 arxiv

Recent vision-language model (VLM)-based approaches have achieved impressive results on SVG generation. However, because they generate only text and lack visual signals during decoding, they often struggle with complex s…

Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations

2025-06-23 · Jiaming Han, Hao Chen, Yang Zhao, Hanyu Wang 외

This paper presents a multimodal framework that attempts to unify visual understanding and generation within a shared discrete semantic representation. At its core is the Text-Aligned Tokenizer (TA-Tok), which converts i…

TAR

BrainCLIP: Bridging Brain and Visual-Linguistic Representation Via CLIP for Generic Natural Visual Stimulus Decoding

2023-02-25 · Yulong Liu, Yongqiang Ma, Wei Zhou, Guibo Zhu 외

Due to the lack of paired samples and the low signal-to-noise ratio of functional MRI (fMRI) signals, reconstructing perceived natural images or decoding their semantic contents from fMRI data are challenging tasks. In t…

Brain DecodingImage GenerationImage ReconstructionImage-text matching+1