paper-with-me

MusicCaps

홈페이지 · 논문 84편

MusicCaps is a dataset composed of 5.5k music-text pairs, with rich text descriptions provided by human experts. For each 10-second music clip, MusicCaps provides: 1) A free-text caption consisting of four sentences on average, describing the music and 2) A list of music aspects, describing genre, mood, tempo, singer voices, instrumentation, dissonances, rhythm, etc. Source:MusicLM: Generating Music From Text

TextsMusic EnglishVietnamese

벤치마크

Text-to-Music Generation on MusicCaps 결과 21개