paper-with-me

Papers

End-to-end Lyrics Alignment for Polyphonic Music Using an Audio-to-Character Recognition Model

2019-02-18 · Daniel Stoller, Simon Durand, Sebastian Ewert

Time-aligned lyrics can enrich the music listening experience by enabling karaoke, text-based song retrieval and intra-song navigation, and other applications. Compared to text-to-speech alignment, lyrics alignment remains highly challenging, despite many attempts to combine numerous sub-modules including vocal separation and detection in an effort to break down the problem. Furthermore, training required fine-grained annotations to be available in some form. Here, we present a novel system based on a modified Wave-U-Net architecture, which predicts character probabilities directly from raw audio using learnt multi-scale representations of the various signal components. There are no sub-modules whose interdependencies need to be optimized. Our training procedure is designed to work with weak, line-level annotations available in the real world. With a mean alignment error of 0.35s on a standard dataset our system outperforms the state-of-the-art by an order of magnitude.

📄 PDF Abstract BibTeX arXiv:1902.06797

Code (2)

f90/jamendolyrics 공식 구현
guxm2021/alt_speechbrain pytorch

Tasks

Retrievaltext-to-speechText to Speech

Similar Papers 제목 키워드 기반

Acoustic Modeling for Automatic Lyrics-to-Audio Alignment

2019-06-25 · Chitralekha Gupta, Emre Yilmaz, Haizhou Li

Automatic lyrics to polyphonic audio alignment is a challenging task not only because the vocals are corrupted by background music, but also there is a lack of annotated polyphonic corpus for effective acoustic modeling.…

Adapting pretrained speech model for Mandarin lyrics transcription and alignment

2023-11-21 · Jun-You Wang, Chon-In Leong, Yu-Chen Lin, Li Su 외

The tasks of automatic lyrics transcription and lyrics alignment have witnessed significant performance improvements in the past few years. However, most of the previous works only focus on English in which large-scale d…

Automatic Lyrics TranscriptionData Augmentation

Genre-conditioned Acoustic Models for Automatic Lyrics Transcription of Polyphonic Music

2022-04-07 · Xiaoxue Gao, Chitralekha Gupta, Haizhou Li

Lyrics transcription of polyphonic music is challenging not only because the singing vocals are corrupted by the background music, but also because the background music and the singing style vary across music genres, suc…

Automatic Lyrics Transcription

PoLyScriber: Integrated Fine-tuning of Extractor and Lyrics Transcriber for Polyphonic Music

2022-07-15 · Xiaoxue Gao, Chitralekha Gupta, Haizhou Li

Lyrics transcription of polyphonic music is challenging as the background music affects lyrics intelligibility. Typically, lyrics transcription can be performed by a two-step pipeline, i.e. a singing vocal extraction fro…

Music-robust Automatic Lyrics Transcription of Polyphonic Music

2022-04-07 · Xiaoxue Gao, Chitralekha Gupta, Haizhou Li

Lyrics transcription of polyphonic music is challenging because singing vocals are corrupted by the background music. To improve the robustness of lyrics transcription to the background music, we propose a strategy of co…

Automatic Lyrics TranscriptionLanguage ModelingLanguage Modelling