paper-with-me

홈 › Papers

Latent Image Animator: Learning to Animate Images via Latent Space Navigation

2022-03-17 · Yaohui Wang, Di Yang, Francois Bremond, Antitza Dantcheva

Due to the remarkable progress of deep generative models, animating images has become increasingly efficient, whereas associated results have become increasingly realistic. Current animation-approaches commonly exploit structure representation extracted from driving videos. Such structure representation is instrumental in transferring motion from driving videos to still images. However, such approaches fail in case the source image and driving video encompass large appearance variation. Moreover, the extraction of structure information requires additional modules that endow the animation-model with increased complexity. Deviating from such models, we here introduce the Latent Image Animator (LIA), a self-supervised autoencoder that evades need for structure representation. LIA is streamlined to animate images by linear navigation in the latent space. Specifically, motion in generated video is constructed by linear displacement of codes in the latent space. Towards this, we learn a set of orthogonal motion directions simultaneously, and use their linear combination, in order to represent any displacement in the latent space. Extensive quantitative and qualitative analysis suggests that our model systematically and significantly outperforms state-of-art methods on VoxCeleb, Taichi and TED-talk datasets w.r.t. generated quality.

📄 PDF Abstract BibTeX arXiv:2203.09043

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Latent Image Animator: Learning to animate image via latent space navigation

2021-09-29 · ICLR 2022 4 · Yaohui Wang, Di Yang, Francois Bremond, Antitza Dantcheva

Animating images has become increasingly realistic, as well as efficient due to the remarkable progress of Generative Adversarial Networks (GANs) and auto-encoder. Current animation-approaches commonly exploit structure …

Image AnimationVideo Generation

AnimateZero: Video Diffusion Models are Zero-Shot Image Animators

2023-12-06 · Jiwen Yu, Xiaodong Cun, Chenyang Qi, Yong Zhang 외

Large-scale text-to-video (T2V) diffusion models have great progress in recent years in terms of visual quality, motion and temporal consistency. However, the generation process is still a black box, where all attributes…

Image AnimationVideo Generation

ID-Animator: Zero-Shot Identity-Preserving Human Video Generation

2024-04-23 · Xuanhua He, Quande Liu, Shengju Qian, Xin Wang 외

Generating high-fidelity human video with specified identities has attracted significant attention in the content generation community. However, existing techniques struggle to strike a balance between training efficienc…

AttributeVideo Generation

Latent Painter

2023-08-31 · Shih-Chieh Su

Latent diffusers revolutionized the generative AI and inspired creative art. When denoising the latent, the predicted original image at each step collectively animates the formation. However, the animation is limited by …

Denoising

PE2LGP Animator: A Tool To Animate A Portuguese Sign Language Avatar

2020-05-01 · LREC 2020 5 · Pedro Cabral, Matilde Gon{\c{c}}alves, Hugo Nicolau, Lu{\'\i}sa Coheur 외

Software for the production of sign languages is much less common than for spoken languages. Such software usually relies on 3D humanoid avatars to produce signs which, inevitably, necessitates the use of animation. One …