W-Net : One-Shot Arbitrary-StyleChinese Character Generationwith Deep Neural Networks
Abstract. Due to the huge category number, the sophisticated com-binations of various strokes and radicals, and the free writing or print-ing styles, generating Chinese characters with diverse styles is alwaysconsidered as a difficult task. In this paper, an efficient and general-ized deep framework, namely, the W-Net, is introduced for the one-shotarbitrary-style Chinese character generation task. Specifically, given asingle character (one-shot) with a specific style (e.g., a printed font orhand-writing style), the proposed W-Net model is capable of learningand generating any arbitrary characters sharing the style similar to thegiven single character. Such appealing property was rarely seen in theliterature. We have compared the proposed W-Net framework to manyother competitive methods. Experimental results showed the proposedmethod is significantly superior in the one-shot setting.
Code (1)
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
W-Net: One-Shot Arbitrary-Style Chinese Character Generation with Deep Neural Networks
Due to the huge category number, the sophisticated combinations of various strokes and radicals, and the free writing or printing styles, generating Chinese characters with diverse styles is always considered as a diffic…
FilmWeaver: Weaving Consistent Multi-Shot Videos with Cache-Guided Autoregressive Diffusion
Current video generation models perform well at single-shot synthesis but struggle with multi-shot videos, facing critical challenges in maintaining character and background consistency across shots and flexibly generati…
Video GenerationHierarchical Cross-Modal Talking Face Generationwith Dynamic Pixel-Wise Loss
We devise a cascade GAN approach to generate talking face video, which is robust to different face shapes, view angles, facial characteristics, and noisy audio conditions. Instead of learning a direct mapping from audio …
Improve Variational Autoencoder for Text Generationwith Discrete Latent Bottleneck
Variational autoencoders (VAEs) are essential tools in end-to-end representation learning. However, the sequential text generation common pitfall with VAEs is that the model tends to ignore latent variables with a strong…
DecoderLanguage ModelingLanguage ModellingMachine Translation+6StyleLipSync: Style-based Personalized Lip-sync Video Generation
In this paper, we present StyleLipSync, a style-based personalized lip-sync video generative model that can generate identity-agnostic lip-synchronizing video from arbitrary audio. To generate a video of arbitrary identi…
Video Generation