paper-with-me

Papers

CHULA TTS: A Modularized Text-To-Speech Framework

2014-12-01 · PACLIC 2014 12 · Natthawut Kertkeidkachorn, Supadaech Chanjaradwichai, Proadpran Punyabukkana, Atiwong Suchato
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

text-to-speechText to Speech

Similar Papers 제목 키워드 기반

SpeechNet: A Universal Modularized Model for Speech Processing Tasks

2021-05-07 · Yi-Chen Chen, Po-Han Chi, Shu-wen Yang, Kai-Wei Chang 외

There is a wide variety of speech processing tasks ranging from extracting content information from speech signals to generating speech signals. For different tasks, model networks are usually designed and tuned separate…

Multi-Task Learning

ThaiCoref: Thai Coreference Resolution Dataset

2024-06-10 · Pontakorn Trakuekul, Wei Qi Leong, Charin Polpanumas, Jitkapat Sawatphol 외

While coreference resolution is a well-established research area in Natural Language Processing (NLP), research focusing on Thai language remains limited due to the lack of large annotated corpora. In this work, we intro…

coreference-resolutionCoreference ResolutionCross-Lingual Transfer

ParrotTTS: Text-to-Speech synthesis by exploiting self-supervised representations

2023-03-01 · Neil Shah, Saiteja Kosgi, Vishal Tambrahalli, Neha Sahipjohn 외

We present ParrotTTS, a modularized text-to-speech synthesis model leveraging disentangled self-supervised speech representations. It can train a multi-speaker variant effectively using transcripts from a single speaker.…

Self-Supervised LearningSpeech Synthesistext-to-speechText to Speech+1

aw_nas: A Modularized and Extensible NAS framework

2020-11-25 · Xuefei Ning, Changcheng Tang, Wenshuo Li, Songyi Yang 외

Neural Architecture Search (NAS) has received extensive attention due to its capability to discover neural network architectures in an automated manner. aw_nas is an open-source Python framework implementing various NAS …

Adversarial RobustnessNeural Architecture Search

MParrotTTS: Multilingual Multi-speaker Text to Speech Synthesis in Low Resource Setting

2023-05-19 · Neil Shah, Vishal Tambrahalli, Saiteja Kosgi, Niranjan Pedanekar 외

We present MParrotTTS, a unified multilingual, multi-speaker text-to-speech (TTS) synthesis model that can produce high-quality speech. Benefiting from a modularized training paradigm exploiting self-supervised speech re…

Speech Synthesistext-to-speechText to SpeechText-To-Speech Synthesis