paper-with-me

Papers

Deep Learning the EEG Manifold for Phonological Categorization from Active Thoughts

2019-04-08 · Pramit Saha, Muhammad Abdul-Mageed, Sidney Fels

Speech-related Brain Computer Interfaces (BCI) aim primarily at finding an alternative vocal communication pathway for people with speaking disabilities. As a step towards full decoding of imagined speech from active thoughts, we present a BCI system for subject-independent classification of phonological categories exploiting a novel deep learning based hierarchical feature extraction scheme. To better capture the complex representation of high-dimensional electroencephalography (EEG) data, we compute the joint variability of EEG electrodes into a channel cross-covariance matrix. We then extract the spatio-temporal information encoded within the matrix using a mixed deep neural network strategy. Our model framework is composed of a convolutional neural network (CNN), a long-short term network (LSTM), and a deep autoencoder. We train the individual networks hierarchically, feeding their combined outputs in a final gradient boosting classification step. Our best models achieve an average accuracy of 77.9% across five different binary classification tasks, providing a significant 22.5% improvement over previous methods. As we also show visually, our work demonstrates that the speech imagery EEG possesses significant discriminative information about the intended articulatory movements responsible for natural speech synthesis.

📄 PDF Abstract BibTeX arXiv:1904.04358

Code (0)

등록된 구현이 없습니다.

Tasks

Binary ClassificationDeep LearningEEGElectroencephalogram (EEG)General ClassificationSpeech Synthesis

Similar Papers 제목 키워드 기반

SPEAK YOUR MIND! Towards Imagined Speech Recognition With Hierarchical Deep Learning

2019-04-08 · Pramit Saha, Muhammad Abdul-Mageed, Sidney Fels

Speech-related Brain Computer Interface (BCI) technologies provide effective vocal communication strategies for controlling devices through speech commands interpreted from brain signals. In order to infer imagined speec…

Brain Computer InterfaceGeneral Classificationspeech-recognitionSpeech Recognition+1

Pre-frontal cortex guides dimension-reducing transformations in the occipito-ventral pathway for categorization behaviors

2022-05-09 · Y. Duan, J. Zhan, J. Gross, R. A. A. Ince 외

To interpret our surroundings, the brain uses a visual categorization process. Current theories and models suggest that this process comprises a hierarchy of different computations that transforms complex, high-dimension…

feature selection

Perceptual compensation for tonal context in self-supervised speech models

2026-06-16 · James Kirby, Ioana Krehan, Michele Gubian arxiv

This study examines the extent to which the wav2vec2.0 architecture exhibits evidence of compensation for phonological context. We conducted a pseudo-replication of a perceptional compensation experiment on Mandarin Chin…

A New Manifold Distance Measure for Visual Object Categorization

2016-05-12 · Fengfu Li, Xiayuan Huang, Hong Qiao, Bo Zhang

Manifold distances are very effective tools for visual object recognition. However, most of the traditional manifold distances between images are based on the pixel-level comparison and thus easily affected by image rota…

ClusteringObjectObject CategorizationObject Recognition+1

[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic

2026-02-21 · Kwanghee Choi, Eunjung Yeo, Cheol Jun Cho, David Harwath 외 arxiv

Self-supervised speech models (S3Ms) are known to encode rich phonetic information, yet how this information is structured remains underexplored. We conduct a comprehensive study across 96 languages to analyze the underl…