paper-with-me

Papers

DynaMind: Reconstructing Dynamic Visual Scenes from EEG by Aligning Temporal Dynamics and Multimodal Semantics to Guided Diffusion

2025-09-01 · Junxiang Liu, Junming Lin, Jiangtong Li, Jie Li arxiv

Reconstruction dynamic visual scenes from electroencephalography (EEG) signals remains a primary challenge in brain decoding, limited by the low spatial resolution of EEG, a temporal mismatch between neural recordings and video dynamics, and the insufficient use of semantic information within brain activity. Therefore, existing methods often inadequately resolve both the dynamic coherence and the complex semantic context of the perceived visual stimuli. To overcome these limitations, we introduce DynaMind, a novel framework that reconstructs video by jointly modeling neural dynamics and semantic features via three core modules: a Regional-aware Semantic Mapper (RSM), a Temporal-aware Dynamic Aligner (TDA), and a Dual-Guidance Video Reconstructor (DGVR). The RSM first utilizes a regional-aware encoder to extract multimodal semantic features from EEG signals across distinct brain regions, aggregating them into a unified diffusion prior. In the mean time, the TDA generates a dynamic latent sequence, or blueprint, to enforce temporal consistency between the feature representations and the original neural recordings. Together, guided by the semantic diffusion prior, the DGVR translates the temporal-aware blueprint into a high-fidelity video reconstruction. On the SEED-DV dataset, DynaMind sets a new state-of-the-art (SOTA), boosting reconstructed video accuracies (video- and frame-based) by 12.5 and 10.3 percentage points, respectively. It also achieves a leap in pixel-level quality, showing exceptional visual fidelity and temporal coherence with a 9.4% SSIM improvement and a 19.7% FVMD reduction. This marks a critical advancement, bridging the gap between neural dynamics and high-fidelity visual semantics.

📄 PDF Abstract BibTeX arXiv:2509.01177

Code (0)

등록된 구현이 없습니다.

Tasks

Video ReconstructionBrain Decoding

Similar Papers 제목 키워드 기반

From Static to Dynamic: A Continual Learning Framework for Large Language Models

2023-10-22 · Mingzhe Du, Anh Tuan Luu, Bin Ji, See-Kiong Ng

The vast number of parameters in large language models (LLMs) endows them with remarkable capabilities, allowing them to excel in a variety of natural language processing tasks. However, this complexity also presents cha…

Continual Learning

Retina-Like Visual Image Reconstruction via Spiking Neural Model

2020-06-01 · CVPR 2020 6 · Lin Zhu, Siwei Dong, Jianing Li, Tiejun Huang 외

The high-sensitivity vision of primates, including humans, is mediated by a small retinal region called the fovea. As a novel bio-inspired vision sensor, spike camera mimics the fovea to record the nature scenes by conti…

Image Reconstruction

Aligning Neuronal Coding of Dynamic Visual Scenes with Foundation Vision Models

2024-07-15 · Rining Wu, Feixiang Zhou, Ziwei Yin, Jian K. Liu

Our brains represent the ever-changing environment with neurons in a highly dynamic fashion. The temporal features of visual pixels in dynamic natural scenes are entrapped in the neuronal responses of the retina. It is c…

EvDNeRF: Reconstructing Event Data with Dynamic Neural Radiance Fields

2023-10-03 · Anish Bhattacharya, Ratnesh Madaan, Fernando Cladera, Sai Vemprala 외

We present EvDNeRF, a pipeline for generating event data and training an event-based dynamic NeRF, for the purpose of faithfully reconstructing eventstreams on scenes with rigid and non-rigid deformations that may be too…

NeRF

DynPL-SVO: A Robust Stereo Visual Odometry for Dynamic Scenes

2022-05-17 · Baosheng Zhang, Xiaoguang Ma, Hongjun Ma, Chunbo Luo

Most feature-based stereo visual odometry (SVO) approaches estimate the motion of mobile robots by matching and tracking point features along a sequence of stereo images. However, in dynamic scenes mainly comprising movi…

Motion EstimationVisual Odometry