paper-with-me

Papers

Arbitrary Motion Style Transfer with Multi-condition Motion Latent Diffusion Model

2024-01-01 · CVPR 2024 1 · Wenfeng Song, Xingliang Jin, Shuai Li, Chenglizhao Chen, Aimin Hao, Xia Hou, Ning li, Hong Qin

Computer animation's quest to bridge content and style has historically been a challenging venture with previous efforts often leaning toward one at the expense of the other. This paper tackles the inherent challenge of content-style duality ensuring a harmonious fusion where the core narrative of the content is both preserved and elevated through stylistic enhancements. We propose a novel Multi-condition Motion Latent Diffusion Model (MCM-LDM) for Arbitrary Motion Style Transfer (AMST). Our MCM-LDM significantly emphasizes preserving trajectories recognizing their fundamental role in defining the essence and fluidity of motion content. Our MCM-LDM's cornerstone lies in its ability first to disentangle and then intricately weave together motion's tripartite components: motion trajectory motion content and motion style. The critical insight of MCM-LDM is to embed multiple conditions with distinct priorities. The content channel serves as the primary flow guiding the overall structure and movement while the trajectory and style channels act as auxiliary components and synchronize with the primary one dynamically. This mechanism ensures that multi-conditions can seamlessly integrate into the main flow enhancing the overall animation without overshadowing the core content. Empirical evaluations underscore the model's proficiency in achieving fluid and authentic motion style transfers setting a new benchmark in the realm of computer animation. The source code and model are available at https://github.com/XingliangJin/MCM-LDM.git.

📄 PDF Abstract BibTeX

Code (1)

xingliangjin/mcm-ldm 공식 구현 pytorch

Tasks

Motion Style TransferStyle Transfer

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.

Similar Papers 제목 키워드 기반

Motion Puzzle: Arbitrary Motion Style Transfer by Body Part

2022-02-10 · Deok-Kyeong Jang, Soomin Park, Sung-Hee Lee

This paper presents Motion Puzzle, a novel motion style transfer network that advances the state-of-the-art in several important respects. The Motion Puzzle is the first that can control the motion style of individual bo…

Motion GenerationMotion Style TransferMotion SynthesisStyle Transfer

CLIP3Dstyler: Language Guided 3D Arbitrary Neural Style Transfer

2023-05-25 · Ming Gao, Yanwu Xu, Yang Zhao, Tingbo Hou 외

In this paper, we propose a novel language-guided 3D arbitrary neural style transfer method (CLIP3Dstyler). We aim at stylizing any 3D scene with an arbitrary style from a text description, and synthesizing the novel sty…

Style Transfer

AStF: Motion Style Transfer via Adaptive Statistics Fusor

2025-11-06 · Hanmo Chen, Chenghao Xu, Jiexi Yan, Cheng Deng arxiv

Human motion style transfer allows characters to appear less rigidity and more realism with specific style. Traditional arbitrary image style transfer typically process mean and variance which is proved effective. Meanwh…

Style Transfer

Parameter-Free Style Projection for Arbitrary Style Transfer

2020-03-17 · Siyu Huang, Haoyi Xiong, Tianyang Wang, Bihan Wen 외

Arbitrary image style transfer is a challenging task which aims to stylize a content image conditioned on arbitrary style images. In this task the feature-level content-style transformation plays a vital role for proper …

Style Transfer

High-fidelity Generalized Emotional Talking Face Generation with Multi-modal Emotion Space Learning

2023-05-04 · CVPR 2023 1 · Chao Xu, Junwei Zhu, Jiangning Zhang, Yue Han 외

Recently, emotional talking face generation has received considerable attention. However, existing methods only adopt one-hot coding, image, or audio as emotion conditions, thus lacking flexible control in practical appl…

Face GenerationTalking Face Generation