paper-with-me

홈 › Papers

A Survey on Recent Deep Learning-driven Singing Voice Synthesis Systems

2021-10-06 · Yin-Ping Cho, Fu-Rong Yang, Yung-Chuan Chang, Ching-Ting Cheng, Xiao-Han Wang, Yi-Wen Liu

Singing voice synthesis (SVS) is a task that aims to generate audio signals according to musical scores and lyrics. With its multifaceted nature concerning music and language, producing singing voices indistinguishable from that of human singers has always remained an unfulfilled pursuit. Nonetheless, the advancements of deep learning techniques have brought about a substantial leap in the quality and naturalness of synthesized singing voice. This paper aims to review some of the state-of-the-art deep learning-driven SVS systems. We intend to summarize their deployed model architectures and identify the strengths and limitations for each of the introduced systems. Thereby, we picture the recent advancement trajectory of this field and conclude the challenges left to be resolved both in commercial applications and academic research.

📄 PDF Abstract BibTeX arXiv:2110.02511

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningSinging Voice Synthesis

Similar Papers 제목 키워드 기반

Controllable Singing Voice Synthesis using Phoneme-Level Energy Sequence

2025-09-08 · Yerin Ryu, Inseop Shin, Chanwoo Kim arxiv

Controllable Singing Voice Synthesis (SVS) aims to generate expressive singing voices reflecting user intent. While recent SVS systems achieve high audio quality, most rely on probabilistic modeling, limiting precise con…

ConSinger: Efficient High-Fidelity Singing Voice Generation with Minimal Steps

2024-10-20 · Yulin Song, Guorui Sang, Jing Yu, Chuangbai Xiao

Singing voice synthesis (SVS) system is expected to generate high-fidelity singing voice from given music scores (lyrics, duration and pitch). Recently, diffusion models have performed well in this field. However, sacrif…

Singing Voice Synthesis

A Melody-Unsupervision Model for Singing Voice Synthesis

2021-10-13 · Soonbeom Choi, Juhan Nam

Recent studies in singing voice synthesis have achieved high-quality results leveraging advances in text-to-speech models based on deep neural networks. One of the main issues in training singing voice synthesis models i…

modelSinging Voice Synthesistext-to-speechText to Speech

MLP Singer: Towards Rapid Parallel Singing Voice Synthesis

2021-06-15 · arXiv 2021 6 · Jaesung Tae, Hyeongju Kim, Younggun Lee

Recent developments in deep learning have significantly improved the quality of synthesized singing voice audio. However, prominent neural singing voice synthesis systems suffer from slow inference speed due to their aut…

image-classificationSinging Voice Synthesis

Towards Improving the Expressiveness of Singing Voice Synthesis with BERT Derived Semantic Information

2023-08-31 · Shaohuan Zhou, Shun Lei, Weiya You, Deyi Tuo 외

This paper presents an end-to-end high-quality singing voice synthesis (SVS) system that uses bidirectional encoder representation from Transformers (BERT) derived semantic embeddings to improve the expressiveness of the…

Singing Voice Synthesis