paper-with-me

홈 › Papers

Deep Image Synthesis from Intuitive User Input: A Review and Perspectives

2021-07-09 · Yuan Xue, Yuan-Chen Guo, Han Zhang, Tao Xu, Song-Hai Zhang, Xiaolei Huang

In many applications of computer graphics, art and design, it is desirable for a user to provide intuitive non-image input, such as text, sketch, stroke, graph or layout, and have a computer system automatically generate photo-realistic images that adhere to the input content. While classic works that allow such automatic image content generation have followed a framework of image retrieval and composition, recent advances in deep generative models such as generative adversarial networks (GANs), variational autoencoders (VAEs), and flow-based methods have enabled more powerful and versatile image generation tasks. This paper reviews recent works for image synthesis given intuitive user input, covering advances in input versatility, image generation methodology, benchmark datasets, and evaluation metrics. This motivates new perspectives on input representation and interactivity, cross pollination between major image generation paradigms, and evaluation and comparison of generation methods.

📄 PDF Abstract BibTeX arXiv:2107.04240

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationImage RetrievalRetrieval

Similar Papers 제목 키워드 기반

Editable Image Elements for Controllable Synthesis

2024-04-24 · Jiteng Mu, Michaël Gharbi, Richard Zhang, Eli Shechtman 외

Diffusion models have made significant advances in text-guided synthesis tasks. However, editing user-provided images remains challenging, as the high dimensional noise input space of diffusion models is not naturally su…

Co-occurrence Based Texture Synthesis

2020-05-17 · Anna Darzi, Itai Lang, Ashutosh Taklikar, Hadar Averbuch-Elor 외

As image generation techniques mature, there is a growing interest in explainable representations that are easy to understand and intuitive to manipulate. In this work, we turn to co-occurrence statistics, which have lon…

Generative Adversarial NetworkImage GenerationMORPHTexture Classification+1

Review of end-to-end speech synthesis technology based on deep learning

2021-04-20 · Zhaoxi Mu, Xinyu Yang, Yizhuo Dong

As an indispensable part of modern human-computer interaction system, speech synthesis technology helps users get the output of intelligent machine more easily and intuitively, thus has attracted more and more attention.…

Speech Synthesis

Why is "Problems" Predictive of Positive Sentiment? A Case Study of Explaining Unintuitive Features in Sentiment Classification

2024-06-05 · Jiaming Qu, Jaime Arguello, Yue Wang

Explainable AI (XAI) algorithms aim to help users understand how a machine learning model makes predictions. To this end, many approaches explain which input features are most predictive of a target label. However, such …

Sentiment AnalysisSentiment Classification

Adversarial Text-to-Image Synthesis: A Review

2021-01-25 · Stanislav Frolov, Tobias Hinz, Federico Raue, Jörn Hees 외

With the advent of generative adversarial networks, synthesizing images from textual descriptions has recently become an active research area. It is a flexible and intuitive way for conditional image generation with sign…

Adversarial TextConditional Image GenerationDiversityImage Generation