paper-with-me

Papers

VToonify: Controllable High-Resolution Portrait Video Style Transfer

2022-09-22 · Shuai Yang, Liming Jiang, Ziwei Liu, Chen Change Loy

Generating high-quality artistic portrait videos is an important and desirable task in computer graphics and vision. Although a series of successful portrait image toonification models built upon the powerful StyleGAN have been proposed, these image-oriented methods have obvious limitations when applied to videos, such as the fixed frame size, the requirement of face alignment, missing non-facial details and temporal inconsistency. In this work, we investigate the challenging controllable high-resolution portrait video style transfer by introducing a novel VToonify framework. Specifically, VToonify leverages the mid- and high-resolution layers of StyleGAN to render high-quality artistic portraits based on the multi-scale content features extracted by an encoder to better preserve the frame details. The resulting fully convolutional architecture accepts non-aligned faces in videos of variable size as input, contributing to complete face regions with natural motions in the output. Our framework is compatible with existing StyleGAN-based image toonification models to extend them to video toonification, and inherits appealing features of these models for flexible style control on color and intensity. This work presents two instantiations of VToonify built upon Toonify and DualStyleGAN for collection-based and exemplar-based portrait video style transfer, respectively. Extensive experimental results demonstrate the effectiveness of our proposed VToonify framework over existing methods in generating high-quality and temporally-coherent artistic portrait videos with flexible style controls.

📄 PDF Abstract BibTeX arXiv:2209.11224

Code (1)

williamyang1991/vtoonify 공식 구현 pytorch

Tasks

Face AlignmentStyle TransferVideo Style TransferVocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
StyleGAN 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Feedforward Network A Feedforward Network, or a Multilayer Perceptron (MLP), is a neural network with solely densely connected layers. This is the classic neural network architecture of the…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
Adaptive Instance Normalization 설명 없음

Similar Papers 제목 키워드 기반

Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

2024-10-10 · Jiahao Cui, Hui Li, Yao Yao, Hao Zhu 외

Recent advances in latent diffusion-based generative models for portrait image animation, such as Hallo, have achieved impressive results in short-duration video synthesis. In this paper, we present updates to Hallo, int…

4kImage AnimationQuantizationVideo Generation

SPACE: Speech-driven Portrait Animation with Controllable Expression

2022-11-17 · ICCV 2023 1 · Siddharth Gururani, Arun Mallya, Ting-Chun Wang, Rafael Valle 외

Animating portraits using speech has received growing attention in recent years, with various creative and practical use cases. An ideal generated video should have good lip sync with the audio, natural facial expression…

Portrait Animation

High-Fidelity Relightable Monocular Portrait Animation with Lighting-Controllable Video Diffusion Model

2025-02-27 · CVPR 2025 1 · Mingtao Guo, Guanyu Xing, Yanli Liu

Relightable portrait animation aims to animate a static reference portrait to match the head movements and expressions of a driving video while adapting to user-specified or reference lighting conditions. Existing portra…

Portrait Animation

Dynamic Neural Portraits

2022-11-25 · Michail Christos Doukas, Stylianos Ploumpis, Stefanos Zafeiriou

We present Dynamic Neural Portraits, a novel approach to the problem of full-head reenactment. Our method generates photo-realistic video portraits by explicitly controlling head pose, facial expressions and eye gaze. Ou…

Image-to-Image TranslationNeRF

MVPortrait: Text-Guided Motion and Emotion Control for Multi-view Vivid Portrait Animation

2025-01-01 · CVPR 2025 1 · Yukang Lin, Hokit Fung, Jianjin Xu, Zeping Ren 외

Recent portrait animation methods have made significant strides in generating realistic lip synchronization. However, they often lack explicit control over head movements and facial expressions, and cannot produce vi…

Portrait AnimationVideo Generation