paper-with-me

Papers

VoxelFormer: Parameter-Efficient Multi-Subject Visual Decoding from fMRI

2025-09-10 · Chenqian Le, Yilin Zhao, Nikasadat Emami, Kushagra Yadav, Xujin "Chris" Liu, Xupeng Chen, Yao Wang arxiv

Recent advances in fMRI-based visual decoding have enabled compelling reconstructions of perceived images. However, most approaches rely on subject-specific training, limiting scalability and practical deployment. We introduce \textbf{VoxelFormer}, a lightweight transformer architecture that enables multi-subject training for visual decoding from fMRI. VoxelFormer integrates a Token Merging Transformer (ToMer) for efficient voxel compression and a query-driven Q-Former that produces fixed-size neural representations aligned with the CLIP image embedding space. Evaluated on the 7T Natural Scenes Dataset, VoxelFormer achieves competitive retrieval performance on subjects included during training with significantly fewer parameters than existing methods. These results highlight token merging and query-based transformers as promising strategies for parameter-efficient neural decoding.

📄 PDF Abstract BibTeX arXiv:2509.09015

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

VoxelFormer: Bird's-Eye-View Feature Generation based on Dual-view Attention for Multi-view 3D Object Detection

2023-04-03 · Zhuoling Li, Chuanrui Zhang, Wei-Chiu Ma, Yipin Zhou 외

In recent years, transformer-based detectors have demonstrated remarkable performance in 2D visual perception tasks. However, their performance in multi-view 3D object detection remains inferior to the state-of-the-art (…

3D Object Detectionobject-detectionObject Detection

CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding

2024-02-14 · Qiongyi Zhou, Changde Du, Shengpei Wang, Huiguang He

The study of decoding visual neural information faces challenges in generalizing single-subject decoding models to multiple subjects, due to individual differences. Moreover, the limited availability of data from a singl…

Representation Learning

Meta-learning In-Context Enables Training-Free Cross Subject Brain Decoding

2026-04-09 · Mu Nan, Muquan Yu, Weijian Mai, Jacob S. Prince 외 arxiv

Visual decoding from brain signals is a key challenge at the intersection of computer vision and neuroscience, requiring methods that bridge neural representations and computational models of vision. A field-wide goal is…

Brain Decoding

MindAdapter: Few-Shot Parameter-Efficient Residual Calibration of Cross-Subject Brain-to-Visual Decoding Models

2026-05-23 · Jiaxiang Liu, Jiawei Du, Xupeng Chen, Guoqi Li 외 arxiv

Cross-subject brain-to-visual decoding remains a core challenge in brain-computer interfaces due to severe inter-individual variability that induces systematic subject-specific functional misalignment. To address this is…

Wills Aligner: Multi-Subject Collaborative Brain Visual Decoding

2024-04-20 · Guangyin Bao, Qi Zhang, Zixuan Gong, Jialei Zhou 외

Decoding visual information from human brain activity has seen remarkable advancements in recent research. However, the diversity in cortical parcellation and fMRI patterns across individuals has prompted the development…

Cross-Modal RetrievalDiversityImage ReconstructionMeta-Learning+1