paper-with-me

홈 › Papers

MuseChat: A Conversational Music Recommendation System for Videos

2023-10-10 · CVPR 2024 1 · Zhikang Dong, Bin Chen, Xiulong Liu, Pawel Polak, Peng Zhang

Music recommendation for videos attracts growing interest in multi-modal research. However, existing systems focus primarily on content compatibility, often ignoring the users' preferences. Their inability to interact with users for further refinements or to provide explanations leads to a less satisfying experience. We address these issues with MuseChat, a first-of-its-kind dialogue-based recommendation system that personalizes music suggestions for videos. Our system consists of two key functionalities with associated modules: recommendation and reasoning. The recommendation module takes a video along with optional information including previous suggested music and user's preference as inputs and retrieves an appropriate music matching the context. The reasoning module, equipped with the power of Large Language Model (Vicuna-7B) and extended to multi-modal inputs, is able to provide reasonable explanation for the recommended music. To evaluate the effectiveness of MuseChat, we build a large-scale dataset, conversational music recommendation for videos, that simulates a two-turn interaction between a user and a recommender based on accurate music track information. Experiment results show that MuseChat achieves significant improvements over existing video-based music retrieval methods as well as offers strong interpretability and interactability.

📄 PDF Abstract BibTeX arXiv:2310.06282

Code (1)

Dongzhikang/MuseChat-dataset 공식 구현

Tasks

Language ModelingLanguage ModellingLarge Language ModelMusic RecommendationRetrieval

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

WeMusic-Agent: Efficient Conversational Music Recommendation via Knowledge Internalization and Agentic Boundary Learning

2025-12-18 · Wendong Bi, Yirong Mao, Xianglong Liu, Kai Tian 외 arxiv

Personalized music recommendation in conversational scenarios usually requires a deep understanding of user preferences and nuanced musical context, yet existing methods often struggle with balancing specialized domain k…

Recommendation Systems

TALKPLAY: Multimodal Music Recommendation with Large Language Models

2025-02-19 · Seungheon Doh, Keunwoo Choi, Juhan Nam

We present TALKPLAY, a novel multimodal music recommendation system that reformulates recommendation as a token generation problem using large language models (LLMs). By leveraging the instruction-following and natural l…

Conversational RecommendationInstruction FollowingLanguage ModelingLanguage Modelling+5

A Human Subject Study of Named Entity Recognition (NER) in Conversational Music Recommendation Queries

2023-03-13 · Elena V. Epure, Romain Hennequin

We conducted a human subject study of named entity recognition on a noisy corpus of conversational music recommendation queries, with many irregular and novel named entities. We evaluated the human NER linguistic behavio…

Music Recommendationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+1

Just Ask for Music (JAM): Multimodal and Personalized Natural Language Music Recommendation

2025-07-21 · Alessandro B. Melchiorre, Elena V. Epure, Shahed Masoudian, Gustavo Escobedo 외 arxiv

Natural language interfaces offer a compelling approach for music recommendation, enabling users to express complex preferences conversationally. While Large Language Models (LLMs) show promise in this direction, their s…

Knowledge Graph Embedding

Talk the Walk: Synthetic Data Generation for Conversational Music Recommendation

2023-01-27 · Megan Leszczynski, Shu Zhang, Ravi Ganti, Krisztian Balog 외

Recommender systems are ubiquitous yet often difficult for users to control, and adjust if recommendation quality is poor. This has motivated conversational recommender systems (CRSs), with control provided through natur…

Language ModellingMusic RecommendationRecommendation SystemsRetrieval+1