paper-with-me

홈 › Papers

整合語者嵌入向量與後置濾波器於提升個人化合成語音之語者相似度 (Incorporating Speaker Embedding and Post-Filter Network for Improving Speaker Similarity of Personalized Speech Synthesis System)

2021-12-01 · IJCLCLP 2021 12 · Sheng-Yao Wang, Yi-chin Huang
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Synthesis

Similar Papers 제목 키워드 기반

Incorporating speaker embedding and post-filter network for improving speaker similarity of personalized speech synthesis system

2021-10-01 · ROCLING 2021 10 · Sheng-Yao Wang, Yi-chin Huang

In recent years, speech synthesis system can generate speech with high speech quality. However, multi-speaker text-to-speech (TTS) system still require large amount of speech data for each target speaker. In this study, …

Speaker VerificationSpeech Synthesistext-to-speechText to Speech+1

Closing the Gap between Single-User and Multi-User VoiceFilter-Lite

2022-02-24 · Rajeev Rikhye, Quan Wang, Qiao Liang, Yanzhang He 외

VoiceFilter-Lite is a speaker-conditioned voice separation model that plays a crucial role in improving speech recognition and speaker verification by suppressing overlapping speech from non-target speakers. However, one…

Speaker Verificationspeech-recognitionSpeech Recognition

Speaker conditioned acoustic modeling for multi-speaker conversational ASR

2021-04-05 · Srikanth Raj Chetupalli, Sriram Ganapathy

In this paper, we propose a novel approach for the transcription of speech conversations with natural speaker overlap, from single channel speech recordings. The proposed model is a combination of a speaker diarization s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speaker-diarizationSpeaker Diarization+2

Quantitative Evidence on Overlooked Aspects of Enrollment Speaker Embeddings for Target Speaker Separation

2022-10-23 · Xiaoyu Liu, Xu Li, Joan Serrà

Single channel target speaker separation (TSS) aims at extracting a speaker's voice from a mixture of multiple talkers given an enrollment utterance of that speaker. A typical deep learning TSS framework consists of an u…

Speaker IdentificationSpeaker Separation

North America Bixby Speaker Diarization System for the VoxCeleb Speaker Recognition Challenge 2021

2021-09-28 · Myungjong Kim, Taeyeon Ki, Aviral Anshu, Vijendra Raj Apsingekar

This paper describes the submission to the speaker diarization track of VoxCeleb Speaker Recognition Challenge 2021 done by North America Bixby Lab of Samsung Research America. Our speaker diarization system consists of …

Clusteringspeaker-diarizationSpeaker DiarizationSpeaker Recognition+1