paper-with-me

Papers

Voice Conversion using Convolutional Neural Networks

2016-10-27 · Shariq Mobin, Joan Bruna

The human auditory system is able to distinguish the vocal source of thousands of speakers, yet not much is known about what features the auditory system uses to do this. Fourier Transforms are capable of capturing the pitch and harmonic structure of the speaker but this alone proves insufficient at identifying speakers uniquely. The remaining structure, often referred to as timbre, is critical to identifying speakers but we understood little about it. In this paper we use recent advances in neural networks in order to manipulate the voice of one speaker into another by transforming not only the pitch of the speaker, but the timbre. We review generative models built with neural networks as well as architectures for creating neural networks that learn analogies. Our preliminary results converting voices from one speaker to another are encouraging.

📄 PDF Abstract BibTeX arXiv:1610.08927

Code (1)

ShariqM/smcnn 공식 구현

Tasks

Voice Conversion

Similar Papers 제목 키워드 기반

Effects of Convolutional Autoencoder Bottleneck Width on StarGAN-based Singing Technique Conversion

2023-08-19 · Tung-Cheng Su, Yung-Chuan Chang, Yi-Wen Liu

Singing technique conversion (STC) refers to the task of converting from one voice technique to another while leaving the original singer identity, melody, and linguistic components intact. Previous STC studies, as well …

Voice Conversion

StarGANv2-VC: A Diverse, Unsupervised, Non-parallel Framework for Natural-Sounding Voice Conversion

2021-07-21 · Yinghao Aaron Li, Ali Zare, Nima Mesgarani

We present an unsupervised non-parallel many-to-many voice conversion (VC) method using a generative adversarial network (GAN) called StarGAN v2. Using a combination of adversarial source classifier loss and perceptual l…

Generative Adversarial Networktext-to-speechText to SpeechVoice Conversion

NVC-Net: End-to-End Adversarial Voice Conversion

2021-06-02 · Bac Nguyen, Fabien Cardinaux

Voice conversion has gained increasing popularity in many applications of speech synthesis. The idea is to change the voice identity from one speaker into another while keeping the linguistic content unchanged. Many voic…

GPUSpeech SynthesisVoice Conversion

Voice Conversion for Stuttered Speech, Instruments, Unseen Languages and Textually Described Voices

2023-10-12 · Matthew Baas, Herman Kamper

Voice conversion aims to convert source speech into a target voice using recordings of the target speaker as a reference. Newer models are producing increasingly realistic output. But what happens when models are fed wit…

Voice Conversion

ConvS2S-VC: Fully convolutional sequence-to-sequence voice conversion

2018-11-05 · Hirokazu Kameoka, Kou Tanaka, Damian Kwasny, Takuhiro Kaneko 외

This paper proposes a voice conversion (VC) method using sequence-to-sequence (seq2seq or S2S) learning, which flexibly converts not only the voice characteristics but also the pitch contour and duration of input speech.…

Speech EnhancementVoice Conversion