paper-with-me

Papers

Secure & Personalized Music-to-Video Generation via CHARCHA

2025-02-03 · Mehul Agarwal, Gauri Agarwal, Santiago Benoit, Andrew Lippman, Jean Oh

Music is a deeply personal experience and our aim is to enhance this with a fully-automated pipeline for personalized music video generation. Our work allows listeners to not just be consumers but co-creators in the music video generation process by creating personalized, consistent and context-driven visuals based on lyrics, rhythm and emotion in the music. The pipeline combines multimodal translation and generation techniques and utilizes low-rank adaptation on listeners' images to create immersive music videos that reflect both the music and the individual. To ensure the ethical use of users' identity, we also introduce CHARCHA (patent pending), a facial identity verification protocol that protects people against unauthorized use of their face while at the same time collecting authorized images from users for personalizing their videos. This paper thus provides a secure and innovative framework for creating deeply personalized music videos.

📄 PDF Abstract BibTeX arXiv:2502.02610

Code (0)

등록된 구현이 없습니다.

Tasks

RhythmVideo Generation

Similar Papers 제목 키워드 기반

Every Image Listens, Every Image Dances: Music-Driven Image Animation

2025-01-30 · Zhikang Dong, Weituo Hao, Ju-Chiang Wang, Peng Zhang 외

Image animation has become a promising area in multimodal research, with a focus on generating videos from reference images. While prior work has largely emphasized generic video generation guided by text, music-driven d…

Image AnimationVideo Generation

A Training-Free Approach for Music Style Transfer with Latent Diffusion Models

2024-11-24 · Sooyoung Kim, Joonwoo Kwon, Heehwan Wang, Shinjae Yoo 외

Music style transfer, while offering exciting possibilities for personalized music generation, often requires extensive training or detailed textual descriptions. This paper introduces a novel training-free approach leve…

Music GenerationMusic Style TransferStyle Transfer

Video Background Music Generation: Dataset, Method and Evaluation

2022-11-21 · ICCV 2023 1 · Le Zhuo, Zhaokai Wang, Baisen Wang, Yue Liao 외

Music is essential when editing videos, but selecting music manually is difficult and time-consuming. Thus, we seek to automatically generate background music tracks given video input. This is a challenging task since it…

Music GenerationRepresentation LearningRetrieval

SingDance: Compositional Zero-Shot Singing-and-Dancing Video Generation with Role-Aware Audio Conditioning

2026-08-17 · Tao Feng, Xu Li, Xiangyang Luo, Ming Wen 외 arxiv

Generating personalized dance videos from a reference image, text prompt, and audio track requires music-conditioned body motion. Singing-and-dancing adds a second requirement: the visible subject must also articulate th…

Video Generation

VMAS: Video-to-Music Generation via Semantic Alignment in Web Music Videos

2024-09-11 · Yan-Bo Lin, Yu Tian, Linjie Yang, Gedas Bertasius 외

We present a framework for learning to generate background music from video inputs. Unlike existing works that rely on symbolic musical annotations, which are limited in quantity and diversity, our method leverages large…

Contrastive LearningMusic Generation