paper-with-me

Papers

MAD: Multi-Alignment MEG-to-Text Decoding

2024-06-03 · Yiqian Yang, Hyejeong Jo, Yiqun Duan, Qiang Zhang, Jinni Zhou, Won Hee Lee, Renjing Xu, Hui Xiong

Deciphering language from brain activity is a crucial task in brain-computer interface (BCI) research. Non-invasive cerebral signaling techniques including electroencephalography (EEG) and magnetoencephalography (MEG) are becoming increasingly popular due to their safety and practicality, avoiding invasive electrode implantation. However, current works under-investigated three points: 1) a predominant focus on EEG with limited exploration of MEG, which provides superior signal quality; 2) poor performance on unseen text, indicating the need for models that can better generalize to diverse linguistic contexts; 3) insufficient integration of information from other modalities, which could potentially constrain our capacity to comprehensively understand the intricate dynamics of brain activity. This study presents a novel approach for translating MEG signals into text using a speech-decoding framework with multiple alignments. Our method is the first to introduce an end-to-end multi-alignment framework for totally unseen text generation directly from MEG signals. We achieve an impressive BLEU-1 score on the $\textit{GWilliams}$ dataset, significantly outperforming the baseline from 5.49 to 10.44 on the BLEU-1 metric. This improvement demonstrates the advancement of our model towards real-world applications and underscores its potential in advancing BCI research. Code is available at $\href{https://github.com/NeuSpeech/MAD-MEG2text}{https://github.com/NeuSpeech/MAD-MEG2text}$.

📄 PDF Abstract BibTeX arXiv:2406.01512

Code (1)

neuspeech/mad-meg2text 공식 구현 pytorch

Tasks

Brain Computer InterfaceEEGText Generation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Through their eyes: multi-subject Brain Decoding with simple alignment techniques

2023-08-01 · Matteo Ferrante, Tommaso Boccato, Nicola Toschi

Previous brain decoding research primarily involves single-subject studies, reconstructing stimuli via fMRI activity from the same subject. Our study aims to introduce a generalization technique for cross-subject brain d…

Brain Decodingregression

Experience-Calibrated Contrastive Decoding for Mitigating Hallucinations in LM-Based Text-to-Speech

2026-08-01 · Chenlin Liu, Minghui Fang, Zhonghao Bi, Zekai Su 외 arxiv

Language model-based text-to-speech (LM-based TTS) remains vulnerable to speech hallucinations that deviate from the target text. Existing mitigation mainly relies on architectural changes or additional training, while d…

Decoupled Attention Network for Text Recognition

2019-12-21 · Tianwei Wang, Yuanzhi Zhu, Lianwen Jin, Canjie Luo 외

Text recognition has attracted considerable research interests because of its various applications. The cutting-edge text recognition methods are based on attention mechanisms. However, most of attention methods usually …

DecoderHandwritten Text RecognitionScene Text Recognition

Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining

2026-07-25 · Rares A. C. Diaconescu, Iulia Slanina, Alina Florea, Andrei B. Trache 외 arxiv

Reducing toxicity is often framed as a global alignment problem, yet perceptions of harmful language are subjective and context-dependent. We present the first comparative evaluation of training-free methods for aligning…

MindAlign: Decoding Inner Speech from fMRI Signals via Multimodal Embedding Alignment under Limited Data

2026-06-15 · Muxuan Liu, Ichiro Kobayashi, Satoshi Nishida arxiv

Decoding inner speech from non-invasive brain signals remains a fundamental challenge due to the absence of overt linguistic output, limited training data, and large inter-subject variability. Existing brain-to-text appr…

Text Generation