paper-with-me

홈 › Papers

Video Reenactment as Inductive Bias for Content-Motion Disentanglement

2021-01-30 · Juan F. Hernández Albarracín, Adín Ramírez Rivera

Independent components within low-dimensional representations are essential inputs in several downstream tasks, and provide explanations over the observed data. Video-based disentangled factors of variation provide low-dimensional representations that can be identified and used to feed task-specific models. We introduce MTC-VAE, a self-supervised motion-transfer VAE model to disentangle motion and content from videos. Unlike previous work on video content-motion disentanglement, we adopt a chunk-wise modeling approach and take advantage of the motion information contained in spatiotemporal neighborhoods. Our model yields independent per-chunk representations that preserve temporal consistency. Hence, we reconstruct whole videos in a single forward-pass. We extend the ELBO's log-likelihood term and include a Blind Reenactment Loss as an inductive bias to leverage motion disentanglement, under the assumption that swapping motion features yields reenactment between two videos. We evaluate our model with recently-proposed disentanglement metrics and show that it outperforms a variety of methods for video motion-content disentanglement. Experiments on video reenactment show the effectiveness of our disentanglement in the input space where our model outperforms the baselines in reconstruction quality and motion alignment.

📄 PDF Abstract BibTeX arXiv:2102.00324

Code (1)

https://gitlab.com/mipl/mtc-vae 공식 구현 tf

Tasks

DisentanglementInductive BiasMotion Disentanglement

Methods 이 논문이 사용한 방법론

USD Coin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Audio-driven Neural Gesture Reenactment with Video Motion Graphs

2022-07-23 · CVPR 2022 1 · Yang Zhou, Jimei Yang, DIngzeyu Li, Jun Saito 외

Human speech is often accompanied by body gestures including arm and hand gestures. We present a method that reenacts a high-quality video with gestures matching a target speech audio. The key idea of our method is to sp…

valid

IMU2Face: Real-time Gesture-driven Facial Reenactment

2017-12-18 · Justus Thies, Michael Zollhöfer, Matthias Nießner

We present IMU2Face, a gesture-driven facial reenactment system. To this end, we combine recent advances in facial motion capture and inertial measurement units (IMUs) to control the facial expressions of a person in a t…

AnyAct: Towards Human Reenactment of Character Motion From Video

2026-05-15 · Liuhan Chen, Lei Zhong, Jiawei Wang, Qin Shuai 외 arxiv

We study the problem of directly deriving an initial human reenactment from a monocular video of a non-human character. Our goal is not to reconstruct the source character itself but to reinterpret its motion as a plausi…

Automatic Face Reenactment

2016-02-08 · CVPR 2014 6 · Pablo Garrido, Levi Valgaerts, Ole Rehmsen, Thorsten Thormaehlen 외

We propose an image-based, facial reenactment system that replaces the face of an actor in an existing target video with the face of a user from a source video, while preserving the original target performance. Our syste…

ClusteringFace ModelFace ReenactmentFace Transfer+2

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

2026-03-16 · Jinguang Tong, Jinbo Wu, Kaisiyuan Wang, Zhelun Shen 외 arxiv

Human-Object Interaction (HOI) video reenactment aims to transfer the interaction dynamics of a source video to a novel target object while preserving realistic hand-object coordination. Existing methods typically rely o…

Video Generation