paper-with-me

홈 › Papers

FISTNet: FusIon of STyle-path generative Networks for Facial Style Transfer

2023-07-18 · Sunder Ali Khowaja, Lewis Nkenyereye, Ghulam Mujtaba, Ik Hyun Lee, Giancarlo Fortino, Kapal Dev

With the surge in emerging technologies such as Metaverse, spatial computing, and generative AI, the application of facial style transfer has gained a lot of interest from researchers as well as startups enthusiasts alike. StyleGAN methods have paved the way for transfer-learning strategies that could reduce the dependency on the huge volume of data that is available for the training process. However, StyleGAN methods have the tendency of overfitting that results in the introduction of artifacts in the facial images. Studies, such as DualStyleGAN, proposed the use of multipath networks but they require the networks to be trained for a specific style rather than generating a fusion of facial styles at once. In this paper, we propose a FusIon of STyles (FIST) network for facial images that leverages pre-trained multipath style transfer networks to eliminate the problem associated with lack of huge data volume in the training phase along with the fusion of multiple styles at the output. We leverage pre-trained styleGAN networks with an external style pass that use residual modulation block instead of a transform coding block. The method also preserves facial structure, identity, and details via the gated mapping unit introduced in this study. The aforementioned components enable us to train the network with very limited amount of data while generating high-quality stylized images. Our training process adapts curriculum learning strategy to perform efficient, flexible style and model fusion in the generative space. We perform extensive experiments to show the superiority of FISTNet in comparison to existing state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2307.09020

Code (0)

등록된 구현이 없습니다.

Tasks

Style TransferTransfer Learning

Methods 이 논문이 사용한 방법론

Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Feedforward Network A Feedforward Network, or a Multilayer Perceptron (MLP), is a neural network with solely densely connected layers. This is the classic neural network architecture of the…
Adaptive Instance Normalization 설명 없음
HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
StyleGAN 설명 없음

Similar Papers 제목 키워드 기반

DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models

2023-09-30 · Zhiyao Sun, Tian Lv, Sheng Ye, Matthieu Lin 외

The generation of stylistic 3D facial animations driven by speech presents a significant challenge as it requires learning a many-to-many mapping between speech, style, and the corresponding natural facial motion. Howeve…

SCHIGAND: A Synthetic Facial Generation Mode Pipeline

2026-01-23 · Ananya Kadali, Sunnie Jehan-Morrison, Orasiki Wellington, Barney Evans 외 arxiv

The growing demand for diverse and high-quality facial datasets for training and testing biometric systems is challenged by privacy regulations, data scarcity, and ethical concerns. Synthetic facial images offer a potent…

Enhancing Quality of Pose-varied Face Restoration with Local Weak Feature Sensing and GAN Prior

2022-05-28 · Kai Hu, Yu Liu, Renhe Liu, Wei Lu 외

Facial semantic guidance (including facial landmarks, facial heatmaps, and facial parsing maps) and facial generative adversarial networks (GAN) prior have been widely used in blind face restoration (BFR) in recent years…

Blind Face RestorationSuper-Resolution

FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing

2024-12-26 · Wanglong Lu, Jikai Wang, Xiaogang Jin, Xianta Jiang 외

Existing facial editing methods have achieved remarkable results, yet they often fall short in supporting multimodal conditional local facial editing. One of the significant evidences is that their output image quality d…

AttributeFacial Editing

MIRRORTALK: Forging Personalized Avatars Via Disentangled Style and Hierarchical Motion Control

2026-01-30 · Renjie Lu, Xulong Zhang, Xiaoyang Qu, Jianzong Wang 외 arxiv

Synthesizing personalized talking faces that uphold and highlight a speaker's unique style while maintaining lip-sync accuracy remains a significant challenge. A primary limitation of existing approaches is the intrinsic…