Deep Learning for Singing Processing: Achievements, Challenges and Impact on Singers and Listeners
This paper summarizes some recent advances on a set of tasks related to the processing of singing using state-of-the-art deep learning techniques. We discuss their achievements in terms of accuracy and sound quality, and the current challenges, such as availability of data and computing resources. We also discuss the impact that these advances do and will have on listeners and singers when they are integrated in commercial applications.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Analysis and Detection of Singing Techniques in Repertoires of J-POP Solo Singers
In this paper, we focus on singing techniques within the scope of music information retrieval research. We investigate how singers use singing techniques using real-world recordings of famous solo singers in Japanese pop…
DescriptiveInformation RetrievalMusic Information RetrievalRetrievalDNN-based ensemble singing voice synthesis with interactions between singers
We propose a singing voice synthesis (SVS) method for a more unified ensemble singing voice by modeling interactions between singers. Most existing SVS methods aim to synthesize a solo voice, and do not consider interact…
Singing Voice SynthesisUnitySinging Voice Graph Modeling for SingFake Detection
Detecting singing voice deepfakes, or SingFake, involves determining the authenticity and copyright of a singing voice. Existing models for speech deepfake detection have struggled to adapt to unseen attacks in this uniq…
DeepFake DetectionFace SwappingRhythmSingFake: Singing Voice Deepfake Detection
The rise of singing voice synthesis presents critical challenges to artists and industry stakeholders over unauthorized voice usage. Unlike synthesized speech, synthesized singing voices are typically released in songs c…
DeepFake DetectionFace SwappingSinging Voice SynthesisSynthetic Speech DetectionGTSinger: A Global Multi-Technique Singing Corpus with Realistic Music Scores for All Singing Tasks
The scarcity of high-quality and multi-task singing datasets significantly hinders the development of diverse controllable and personalized singing tasks, as existing singing datasets suffer from low quality, limited div…
AllSinging Voice SynthesisStyle TransferVocal technique classification