paper-with-me

Papers

Decoding EEG Speech Perception with Transformers and VAE-based Data Augmentation

2025-01-08 · Terrance Yu-Hao Chen, Yulin Chen, Pontus Soederhaell, Sadrishya Agrawal, Kateryna Shapovalenko

Decoding speech from non-invasive brain signals, such as electroencephalography (EEG), has the potential to advance brain-computer interfaces (BCIs), with applications in silent communication and assistive technologies for individuals with speech impairments. However, EEG-based speech decoding faces major challenges, such as noisy data, limited datasets, and poor performance on complex tasks like speech perception. This study attempts to address these challenges by employing variational autoencoders (VAEs) for EEG data augmentation to improve data quality and applying a state-of-the-art (SOTA) sequence-to-sequence deep learning architecture, originally successful in electromyography (EMG) tasks, to EEG-based speech decoding. Additionally, we adapt this architecture for word classification tasks. Using the Brennan dataset, which contains EEG recordings of subjects listening to narrated speech, we preprocess the data and evaluate both classification and sequence-to-sequence models for EEG-to-words/sentences tasks. Our experiments show that VAEs have the potential to reconstruct artificial EEG data for augmentation. Meanwhile, our sequence-to-sequence model achieves more promising performance in generating sentences compared to our classification model, though both remain challenging tasks. These findings lay the groundwork for future research on EEG speech perception decoding, with possible extensions to speech production tasks such as silent or imagined speech.

📄 PDF Abstract BibTeX arXiv:2501.04359

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationEEGElectromyography (EMG)

Similar Papers 제목 키워드 기반

Decoding Part-of-Speech from Human EEG Signals

2022-05-01 · ACL 2022 5 · Alex Murphy, Bernd Bohnet, Ryan Mcdonald, Uta Noppeney

This work explores techniques to predict Part-of-Speech (PoS) tags from neural signals measured at millisecond resolution with electroencephalography (EEG) during text reading. We first show that information about word l…

Data AugmentationEEGElectroencephalogram (EEG)POS+1

Towards unified brain-to-text decoding across speech production and perception

2026-03-13 · Zhizhang Yuan, Yang Yang, Gaorui Zhang, Baowen Cheng 외 arxiv

Speech production and perception are the main ways humans communicate daily. Prior brain-to-text decoding studies have largely focused on a single modality and alphabetic languages. Here, we present a unified brain-to-se…

Transfer Learning for Robust Low-Resource Children's Speech ASR with Transformers and Source-Filter Warping

2022-06-19 · Jenthe Thienpondt, Kris Demuynck

Automatic Speech Recognition (ASR) systems are known to exhibit difficulties when transcribing children's speech. This can mainly be attributed to the absence of large children's speech corpora to train robust ASR models…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+2

MEG-to-MEG Transfer Learning and Cross-Task Speech/Silence Detection with Limited Data

2026-02-20 · Xabier de Zuazo, Vincenzo Verbeni, Eva Navas, Ibon Saratxaga 외 arxiv

Data-efficient neural decoding is a central challenge for speech brain-computer interfaces. We present the first demonstration of transfer learning and cross-task decoding for MEG-based speech models spanning perception …

Transfer Learning

Towards Decoding Brain Activity During Passive Listening of Speech

2024-02-26 · Milán András Fodor, Tamás Gábor Csapó, Frigyes Viktor Arthur

The aim of the study is to investigate the complex mechanisms of speech perception and ultimately decode the electrical changes in the brain accruing while listening to speech. We attempt to decode heard speech from intr…

Brain Computer InterfaceSpeech Synthesis