paper-with-me

Papers

An Algorithm Based on Empirical Methods, for Text-to-Tuneful-Speech Synthesis of Sanskrit Verse

2014-09-15 · Rama N., Meenakshi Lakshmanan

The rendering of Sanskrit poetry from text to speech is a problem that has not been solved before. One reason may be the complications in the language itself. We present unique algorithms based on extensive empirical analysis, to synthesize speech from a given text input of Sanskrit verses. Using a pre-recorded audio units database which is itself tremendously reduced in size compared to the colossal size that would otherwise be required, the algorithms work on producing the best possible, tunefully rendered chanting of the given verse. His would enable the visually impaired and those with reading disabilities to easily access the contents of Sanskrit verses otherwise available only in writing.

📄 PDF Abstract BibTeX arXiv:1409.4169

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Synthesistext-to-speechText to Speech

Similar Papers 제목 키워드 기반

Neural Melody Composition from Lyrics

2018-09-12 · Hangbo Bao, Shaohan Huang, Furu Wei, Lei Cui 외

In this paper, we study a novel task that learns to compose music from natural language. Given the lyrics as input, we propose a melody composition model that generates lyrics-conditional melody as well as the exact alig…

Decoder

Boosting Large Language Model for Speech Synthesis: An Empirical Study

2023-12-30 · Hongkun Hao, Long Zhou, Shujie Liu, Jinyu Li 외

Large language models (LLMs) have made significant advancements in natural language processing and are concurrently extending the language ability to other modalities, such as speech and vision. Nevertheless, most of the…

Language ModelingLanguage ModellingLarge Language ModelSpeech Synthesis+3

LAST: Language Model Aware Speech Tokenization

2024-09-05 · Arnon Turetzky, Yossi Adi

Speech tokenization serves as the foundation of speech language model (LM), enabling them to perform various tasks such as spoken language modeling, text-to-speech, speech-to-text, etc. Most speech tokenizers are trained…

Language ModelingLanguage ModellingmodelQuantization+4

Conditional LSTM-GAN for Melody Generation from Lyrics

2019-08-15 · Yi Yu, Abhishek Srivastava, Simon Canales

Melody generation from lyrics has been a challenging research issue in the field of artificial intelligence and music, which enables to learn and discover latent relationship between interesting lyrics and accompanying m…

Generative Adversarial Network

Enhancing dysarthria speech feature representation with empirical mode decomposition and Walsh-Hadamard transform

2023-12-30 · Ting Zhu, Shufei Duan, Camille Dingam, HuiZhi Liang 외

Dysarthria speech contains the pathological characteristics of vocal tract and vocal fold, but so far, they have not yet been included in traditional acoustic feature sets. Moreover, the nonlinearity and non-stationarity…

imbalanced classification