paper-with-me

Papers

Multi-Level Bidirectional Biomimetic Learning for EEG-Based Visual Decoding

2026-05-06 · Jingtao Liu, Peiliang Gong, Chuhang Zheng, Yiheng Liu, Qi Zhu arxiv

EEG-based visual neural decoding aims to align neural responses with visual stimuli for tasks such as image retrieval. However, limited paired data and a fundamental mismatch between high-fidelity digital images and biological visual perception - distorted by retinotopic mapping and subject-specific neuroanatomy - severely impede cross-modal alignment. To address this, we propose MB2L, a Multi-Level Bidirectional Biomimetic Learning framework that incorporates structured physiological inductive biases into representation learning. Specifically, we propose Adaptive Blur with Visual Priors to mitigate perceptual-structural mismatch by reweighting visual inputs according to retinotopic priors. We further propose Biomimetic Visual Feature Extraction to learn multi-level visual representations consistent with hierarchical cortical processing, enhancing subject-invariant encoding. These modules are jointly optimized via Multi-level Bidirectional Contrastive Learning, which aligns EEG and visual features in a shared semantic space through bidirectional contrastive objectives. Experiments show MB2L achieves 80.5% Top-1 and 97.6% Top-5 accuracy on zero-shot EEG-to-image retrieval, significantly outperforming prior methods and demonstrating strong generalization across subjects and experimental settings.

📄 PDF Abstract BibTeX arXiv:2605.04680

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningContrastive LearningImage Retrieval

Similar Papers 제목 키워드 기반

NeuroFlow: Toward Unified Visual Encoding and Decoding from Neural Activity

2026-04-10 · Weijian Mai, Mu Nan, Yu Zhu, Jiahang Cao 외 arxiv

Visual encoding and decoding models act as gateways to understanding the neural mechanisms underlying human visual perception. Typically, visual encoding models that predict brain activity from stimuli and decoding model…

Computational Efficiency

Bidirectional Beam Search: Forward-Backward Inference in Neural Sequence Models for Fill-in-the-Blank Image Captioning

2017-05-24 · CVPR 2017 7 · Qing Sun, Stefan Lee, Dhruv Batra

We develop the first approximate inference algorithm for 1-Best (and M-Best) decoding in bidirectional neural sequence models by extending Beam Search (BS) to reason about both forward and backward time dependencies. Bea…

Image CaptioningSentence

MLBiNet: A Cross-Sentence Collective Event Detection Network

2021-05-20 · ACL 2021 5 · Dongfang Lou, Zhilin Liao, Shumin Deng, Ningyu Zhang 외

We consider the problem of collectively detecting multiple events, particularly in cross-sentence settings. The key to dealing with the problem is to encode semantic information and model event inter-dependency at a docu…

DecoderEvent DetectionEvent ExtractionSentence+1

Cross-Subject Mind Decoding from Inaccurate Representations

2025-07-25 · Yangyang Xu, Bangzhen Liu, Wenqi Shao, Yong Du 외 arxiv

Decoding stimulus images from fMRI signals has advanced with pre-trained generative models. However, existing methods struggle with cross-subject mappings due to cognitive variability and subject-specific differences. Th…

Diffusion Large Language Models for Visual Speech Recognition

2026-05-27 · Jeong Hun Yeo, Chae Won Kim, Hyeongseop Rha, Yong Man Ro arxiv

Existing Visual Speech Recognition (VSR) systems commonly rely on left-to-right autoregressive decoding, which can force premature decisions on visually ambiguous tokens before sufficient context is available. We propose…

Visual Speech Recognition