paper-with-me

홈 › Papers

Agent-Driven Large Language Models for Mandarin Lyric Generation

2024-10-02 · Hong-Hsiang Liu, Yi-Wen Liu

Generative Large Language Models have shown impressive in-context learning abilities, performing well across various tasks with just a prompt. Previous melody-to-lyric research has been limited by scarce high-quality aligned data and unclear standard for creativeness. Most efforts focused on general themes or emotions, which are less valuable given current language model capabilities. In tonal contour languages like Mandarin, pitch contours are influenced by both melody and tone, leading to variations in lyric-melody fit. Our study, validated by the Mpop600 dataset, confirms that lyricists and melody writers consider this fit during their composition process. In this research, we developed a multi-agent system that decomposes the melody-to-lyric task into sub-tasks, with each agent controlling rhyme, syllable count, lyric-melody alignment, and consistency. Listening tests were conducted via a diffusion-based singing voice synthesizer to evaluate the quality of lyrics generated by different agent groups.

📄 PDF Abstract BibTeX arXiv:2410.01450

Code (0)

등록된 구현이 없습니다.

Tasks

In-Context LearningLanguage ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

Adapting pretrained speech model for Mandarin lyrics transcription and alignment

2023-11-21 · Jun-You Wang, Chon-In Leong, Yu-Chen Lin, Li Su 외

The tasks of automatic lyrics transcription and lyrics alignment have witnessed significant performance improvements in the past few years. However, most of the previous works only focus on English in which large-scale d…

Automatic Lyrics TranscriptionData Augmentation

A Melody-Conditioned Lyrics Language Model

2018-06-01 · NAACL 2018 6 · Kento Watanabe, Yuichiroh Matsubayashi, Satoru Fukayama, Masataka Goto 외

This paper presents a novel, data-driven language model that produces entire lyrics for a given input melody. Previously proposed models for lyrics generation suffer from the inability of capturing the relationship betwe…

Language ModelingLanguage ModellingmodelSentence

FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

2025-01-24 · Kai-Tuo Xu, Feng-Long Xie, Xu Tang, Yao Hu

We present FireRedASR, a family of large-scale automatic speech recognition (ASR) models for Mandarin, designed to meet diverse requirements in superior performance and optimal efficiency across various applications. Fir…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Computational EfficiencyDecoder+3

Automatic Neural Lyrics and Melody Composition

2020-11-12 · Gurunath Reddy Madhumani, Yi Yu, Florian Harscoët, Simon Canales 외

In this paper, we propose a technique to address the most challenging aspect of algorithmic songwriting process, which enables the human community to discover original lyrics, and melodies suitable for the generated lyri…

DecoderSentence

Semantic Communities and Boundary-Spanning Lyrics in K-pop: A Graph-Based Unsupervised Analysis

2026-02-13 · Oktay Karakuş arxiv

Large-scale lyric corpora present unique challenges for data-driven analysis, including the absence of reliable annotations, multilingual content, and high levels of stylistic repetition. Most existing approaches rely on…

Community Detection