paper-with-me

홈 › Papers

Movie Weaver: Tuning-Free Multi-Concept Video Personalization with Anchored Prompts

2025-02-04 · CVPR 2025 1 · Feng Liang, Haoyu Ma, Zecheng He, Tingbo Hou, Ji Hou, Kunpeng Li, Xiaoliang Dai, Felix Juefei-Xu, Samaneh Azadi, Animesh Sinha, Peizhao Zhang, Peter Vajda, Diana Marculescu

Video personalization, which generates customized videos using reference images, has gained significant attention. However, prior methods typically focus on single-concept personalization, limiting broader applications that require multi-concept integration. Attempts to extend these models to multiple concepts often lead to identity blending, which results in composite characters with fused attributes from multiple sources. This challenge arises due to the lack of a mechanism to link each concept with its specific reference image. We address this with anchored prompts, which embed image anchors as unique tokens within text prompts, guiding accurate referencing during generation. Additionally, we introduce concept embeddings to encode the order of reference images. Our approach, Movie Weaver, seamlessly weaves multiple concepts-including face, body, and animal images-into one video, allowing flexible combinations in a single model. The evaluation shows that Movie Weaver outperforms existing methods for multi-concept video personalization in identity preservation and overall quality.

📄 PDF Abstract BibTeX arXiv:2502.07802

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

MasterWeaver: Taming Editability and Face Identity for Personalized Text-to-Image Generation

2024-05-09 · Yuxiang Wei, Zhilong Ji, Jinfeng Bai, Hongzhi Zhang 외

Text-to-image (T2I) diffusion models have shown significant success in personalized text-to-image generation, which aims to generate novel images with human identities indicated by the reference images. Despite promising…

Image GenerationText to Image GenerationText-to-Image Generation

ConceptWeaver: Weaving Disentangled Concepts with Flow

2026-03-30 · Jintao Chen, Aiming Hao, Xiaoqing Chen, Chengyu Bai 외 arxiv

Pre-trained flow-based models excel at synthesizing complex scenes yet lack a direct mechanism for disentangling and customizing their underlying concepts from one-shot real-world sources. To demystify this process, we f…

Beyond Testers' Biases: Guiding Model Testing with Knowledge Bases using LLMs

2023-10-14 · Chenyang Yang, Rishabh Rustogi, Rachel Brower-Sinning, Grace A. Lewis 외

Current model testing work has mostly focused on creating test cases. Identifying what to test is a step that is largely ignored and poorly supported. We propose Weaver, an interactive tool that supports requirements eli…

Stance Detection

SimWeaver: Zero-Shot RGB Sim-to-Real for Deformable Manipulation

2026-06-13 · Wenkang Hu, Haoran Wang, Yitong Li, Liu Liu 외 arxiv

RGB sim-to-real for deformable manipulation has remained largely unsolved without real-world fine-tuning. We present SimWeaver, which trains zero-shot RGB VLA policies on 200 simulated demonstrations per task, reaching a…

Image Generation

Weaver: End-to-End Agentic System Training for Video Interleaved Reasoning

2026-02-05 · Yudi Shi, Shangzhe Di, Qirui Chen, Qinian Wang 외 arxiv

Video reasoning constitutes a comprehensive assessment of a model's capabilities, as it demands robust perceptual and interpretive skills, thereby serving as a means to explore the boundaries of model performance. While …

Reinforcement LearningMultimodal Reasoning