paper-with-me

홈 › Papers

SongSong: A Time Phonograph for Chinese SongCi Music from Thousand of Years Away

2026-02-27 · Jiajia Li, Jiliang Hu, Ziyi Pan, Chong Chen, Zuchao Li, Ping Wang, Lefei Zhang arxiv

Recently, there have been significant advancements in music generation. However, existing models primarily focus on creating modern pop songs, making it challenging to produce ancient music with distinct rhythms and styles, such as ancient Chinese SongCi. In this paper, we introduce SongSong, the first music generation model capable of restoring Chinese SongCi to our knowledge. Our model first predicts the melody from the input SongCi, then separately generates the singing voice and accompaniment based on that melody, and finally combines all elements to create the final piece of music. Additionally, to address the lack of ancient music datasets, we create OpenSongSong, a comprehensive dataset of ancient Chinese SongCi music, featuring 29.9 hours of compositions by various renowned SongCi music masters. To assess SongSong's proficiency in performing SongCi, we randomly select 85 SongCi sentences that were not part of the training set for evaluation against SongSong and music generation platforms such as Suno and SkyMusic. The subjective and objective outcomes indicate that our proposed model achieves leading performance in generating high-quality SongCi music.

📄 PDF Abstract BibTeX arXiv:2602.24071

Code (0)

등록된 구현이 없습니다.

Tasks

Music Generation

Similar Papers 제목 키워드 기반

PoeTone: A Framework for Constrained Generation of Structured Chinese Songci with LLMs

2025-08-04 · Zhan Qu, Shuzhou Yuan, Michael Färber arxiv

This paper presents a systematic investigation into the constrained generation capabilities of large language models (LLMs) in producing Songci, a classical Chinese poetry form characterized by strict structural, tonal, …

SongNet: Rigid Formats Controlled Text Generation

2020-04-17 · ACL 2020 7 · Piji Li, Haisong Zhang, Xiaojiang Liu, Shuming Shi

Neural text generation has made tremendous progress in various tasks. One common characteristic of most of the tasks is that the texts are not restricted to some rigid formats when generating. However, we may confront so…

Language ModellingSentenceText Generation

Large-vocabulary forensic pathological analyses via prototypical cross-modal contrastive learning

2024-07-20 · Chen Shen, Chunfeng Lian, Wanqing Zhang, Fan Wang 외

Forensic pathology is critical in determining the cause and manner of death through post-mortem examinations, both macroscopic and microscopic. The field, however, grapples with issues such as outcome variability, labori…

Contrastive LearningDiagnosticLanguage ModelingLanguage Modelling+1

ChMusic: A Traditional Chinese Music Dataset for Evaluation of Instrument Recognition

2021-08-19 · Xia Gong, Yuxiang Zhu, Haidi Zhu, Haoran Wei

Musical instruments recognition is a widely used application for music information retrieval. As most of previous musical instruments recognition dataset focus on western musical instruments, it is difficult for research…

Information RetrievalInstrument RecognitionMusic Information RetrievalRetrieval

The fusion of phonography and ideographic characters into virtual Chinese characters -- Based on Chinese and English

2024-08-20 · Hongfa Zi, Zhen Liu

The characters used in modern countries are mainly divided into ideographic characters and phonetic characters, both of which have their advantages and disadvantages. Chinese is difficult to learn and easy to master, whi…