paper-with-me

Papers

Applying Phonological Features in Multilingual Text-To-Speech

2021-10-07 · Cong Zhang, Huinan Zeng, Huang Liu, Jiewen Zheng

This study investigates whether phonological features can be applied in text-to-speech systems to generate native and non-native speech in English and Mandarin. We present a mapping of ARPABET/pinyin to SAMPA/SAMPA-SC and then to phonological features. We tested whether this mapping could lead to the successful generation of native, non-native, and code-switched speech in the two languages. We ran two experiments, one with a small dataset and one with a larger dataset. The results proved that phonological features could be used as a feasible input system, although further investigation is needed to improve model performance. The accented output generated by the TTS models also helps with understanding human second language acquisition processes.

📄 PDF Abstract BibTeX arXiv:2110.03609

Code (1)

congzhang365/feature_tts

Tasks

Language Acquisitiontext-to-speechText to Speech

Similar Papers 제목 키워드 기반

Applying Feature Underspecified Lexicon Phonological Features in Multilingual Text-to-Speech

2022-04-14 · Cong Zhang, Huinan Zeng, Huang Liu, Jiewen Zheng

This study investigates whether the phonological features derived from the Featurally Underspecified Lexicon model can be applied in text-to-speech systems to generate native and non-native speech in English and Mandarin…

Language Acquisitiontext-to-speechText to Speech

Multilingual and crosslingual speech recognition using phonological-vector based phone embeddings

2021-07-11 · Chengrui Zhu, Keyu An, Huahuan Zheng, Zhijian Ou

The use of phonological features (PFs) potentially allows language-specific phones to remain linked in training, which is highly desirable for information sharing for multilingual and crosslingual speech recognition meth…

speech-recognitionSpeech Recognition

Multilingual Phonological Feature Recognition with Self-Supervised Speech Models

2026-05-25 · Abner Hernandez, Tomás Arias-Vergara, Daiqi Liu, Andreas Maier 외 arxiv

Phonological features provide a language-general and linguistically grounded representation of speech. We present PhonoQ-2.0, a multilingual frame-level phonological feature recognizer built on self-supervised speech mod…

Phonological Features for 0-shot Multilingual Speech Synthesis

2020-08-06 · Marlene Staib, Tian Huey Teh, Alexandra Torresquintero, Devang S Ram Mohan 외

Code-switching---the intra-utterance use of multiple languages---is prevalent across the world. Within text-to-speech (TTS), multilingual models have been found to enable code-switching. By modifying the linguistic input…

Speech Synthesistext-to-speechText to Speech

Cross-lingual Low Resource Speaker Adaptation Using Phonological Features

2021-11-17 · Georgia Maniati, Nikolaos Ellinas, Konstantinos Markopoulos, Georgios Vamvoukakis 외

The idea of using phonological features instead of phonemes as input to sequence-to-sequence TTS has been recently proposed for zero-shot multilingual speech synthesis. This approach is useful for code-switching, as it f…

Speech Synthesis