paper-with-me

홈 › Papers

Advancing Speech Synthesis using EEG

2020-04-09 · Gautam Krishna, Co Tran, Mason Carnahan, Ahmed Tewfik

In this paper we introduce attention-regression model to demonstrate predicting acoustic features from electroencephalography (EEG) features recorded in parallel with spoken sentences. First we demonstrate predicting acoustic features directly from EEG features using our attention model and then we demonstrate predicting acoustic features from EEG features using a two-step approach where in the first step we use our attention model to predict articulatory features from EEG features and then in second step another attention-regression model is trained to transform the predicted articulatory features to acoustic features. Our proposed attention-regression model demonstrates superior performance compared to the regression model introduced by authors in [1] when tested using their data set for majority of the subjects during test time. The results presented in this paper further advances the work described by authors in [1].

📄 PDF Abstract BibTeX arXiv:2004.04731

Code (0)

등록된 구현이 없습니다.

Tasks

EEGElectroencephalogram (EEG)regressionSpeech Synthesis

Similar Papers 제목 키워드 기반

1000 African Voices: Advancing inclusive multi-speaker multi-accent speech synthesis

2024-06-17 · Sewade Ogun, Abraham T. Owodunni, Tobi Olatunji, Eniola Alese 외

Recent advances in speech synthesis have enabled many useful applications like audio directions in Google Maps, screen readers, and automated content generation on platforms like TikTok. However, these systems are mostly…

DiversitySpeech Synthesis

ARTI-6: Towards Six-dimensional Articulatory Speech Encoding

2025-09-25 · Jihwan Lee, Sean Foley, Thanathai Lertpetchpun, Kevin Huang 외 arxiv

We propose ARTI-6, a compact six-dimensional articulatory speech encoding framework derived from real-time MRI data that captures crucial vocal tract regions including the velum, tongue root, and larynx. ARTI-6 consists …

VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing

2025-11-15 · Zhisheng Zheng, Puyuan Peng, Anuj Diwan, Cong Phuoc Huynh 외 arxiv

We introduce VoiceCraft-X, an autoregressive neural codec language model which unifies multilingual speech editing and zero-shot Text-to-Speech (TTS) synthesis across 11 languages: English, Mandarin, Korean, Japanese, Sp…

Speech Synthesis

Expressivity and Speech Synthesis

2024-04-30 · Andreas Triantafyllopoulos, Björn W. Schuller

Imbuing machines with the ability to talk has been a longtime pursuit of artificial intelligence (AI) research. From the very beginning, the community has not only aimed to synthesise high-fidelity speech that accurately…

Expressive Speech SynthesisSpeech Synthesis

Breeze Taigi: Benchmarks and Models for Taiwanese Hokkien Speech Recognition and Synthesis

2026-02-26 · Yu-Siang Lan, Chia-Sheng Liu, Yi-Chang Chen, Po-Chun Hsu 외 arxiv

Taiwanese Hokkien (Taigi) presents unique opportunities for advancing speech technology methodologies that can generalize to diverse linguistic contexts. We introduce Breeze Taigi, a comprehensive framework centered on s…

Synthetic Data GenerationSpeech Recognition