paper-with-me

Papers

DNN-based Speech Synthesis for Indian Languages from ASCII text

2016-08-18 · Srikanth Ronanki, Siva Reddy, Bajibabu Bollepalli, Simon King

Text-to-Speech synthesis in Indian languages has a seen lot of progress over the decade partly due to the annual Blizzard challenges. These systems assume the text to be written in Devanagari or Dravidian scripts which are nearly phonemic orthography scripts. However, the most common form of computer interaction among Indians is ASCII written transliterated text. Such text is generally noisy with many variations in spelling for the same word. In this paper we evaluate three approaches to synthesize speech from such noisy ASCII text: a naive Uni-Grapheme approach, a Multi-Grapheme approach, and a supervised Grapheme-to-Phoneme (G2P) approach. These methods first convert the ASCII text to a phonetic script, and then learn a Deep Neural Network to synthesize speech from that. We train and test our models on Blizzard Challenge datasets that were transliterated to ASCII using crowdsourcing. Our experiments on Hindi, Tamil and Telugu demonstrate that our models generate speech of competetive quality from ASCII text compared to the speech synthesized from the native scripts. All the accompanying transliterated datasets are released for public access.

📄 PDF Abstract BibTeX arXiv:1608.05374

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Synthesistext-to-speechText to SpeechText-To-Speech Synthesis

Similar Papers 제목 키워드 기반

RASMALAI: Resources for Adaptive Speech Modeling in Indian Languages with Accents and Intonations

2025-05-24 · Ashwin Sankar, Yoach Lacombe, Sherry Thomas, Praveen Srinivasa Varadhan 외

We introduce RASMALAI, a large-scale speech dataset with rich text descriptions, designed to advance controllable and expressive text-to-speech (TTS) synthesis for 23 Indian languages and English. It comprises 13,000 hou…

Expressive Speech SynthesisSpeech Synthesistext-to-speechText to Speech

Everyday Speech in the Indian Subcontinent

2024-10-14 · Utkarsh P

India has 1369 languages of which 22 are official. About 13 different scripts are used to represent these languages. A Common Label Set (CLS) was developed based on phonetics to address the issue of large vocabulary of u…

Speech Synthesis

A Unified Framework for Collecting Text-to-Speech Synthesis Datasets for 22 Indian Languages

2024-10-18 · Sujitha Sathiyamoorthy, N Mohana, Anusha Prakash, Hema A Murthy

The performance of a text-to-speech (TTS) synthesis model depends on various factors, of which the quality of the training data is of utmost importance. Millions of data are collected around the globe for various languag…

Speech Synthesistext-to-speechText to SpeechText-To-Speech Synthesis

Exploring an Inter-Pausal Unit (IPU) based Approach for Indic End-to-End TTS Systems

2024-09-18 · Anusha Prakash, Hema A Murthy

Sentences in Indian languages are generally longer than those in English. Indian languages are also considered to be phrase-based, wherein semantically complete phrases are concatenated to make up sentences. Long utteran…

Sentencetext-to-speechText to Speech

Technology Pipeline for Large Scale Cross-Lingual Dubbing of Lecture Videos into Multiple Indian Languages

2022-11-01 · Anusha Prakash, Arun Kumar, Ashish Seth, Bhagyashree Mukherjee 외

Cross-lingual dubbing of lecture videos requires the transcription of the original audio, correction and removal of disfluencies, domain term discovery, text-to-text translation into the target language, chunking of text…

ChunkingRhythmSpeech Synthesistext-to-speech+2