paper-with-me

홈 › Papers

Hierarchy Composition GAN for High-fidelity Image Synthesis

2019-05-12 · Fangneng Zhan, Jiaxing Huang, Shijian Lu

Despite the rapid progress of generative adversarial networks (GANs) in image synthesis in recent years, the existing image synthesis approaches work in either geometry domain or appearance domain alone which often introduces various synthesis artifacts. This paper presents an innovative Hierarchical Composition GAN (HIC-GAN) that incorporates image synthesis in geometry and appearance domains into an end-to-end trainable network and achieves superior synthesis realism in both domains simultaneously. We design an innovative hierarchical composition mechanism that is capable of learning realistic composition geometry and handling occlusions while multiple foreground objects are involved in image composition. In addition, we introduce a novel attention mask mechanism that guides to adapt the appearance of foreground objects which also helps to provide better training reference for learning in geometry domain. Extensive experiments on scene text image synthesis, portrait editing and indoor rendering tasks show that the proposed HIC-GAN achieves superior synthesis performance qualitatively and quantitatively.

📄 PDF Abstract BibTeX arXiv:1905.04693

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationVocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dogecoin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

StyleT2I: Toward Compositional and High-Fidelity Text-to-Image Synthesis

2022-03-29 · CVPR 2022 1 · Zhiheng Li, Martin Renqiang Min, Kai Li, Chenliang Xu

Although progress has been made for text-to-image synthesis, previous methods fall short of generalizing to unseen or underrepresented attribute compositions in the input text. Lacking compositionality could have severe …

AttributeFairnessImage GenerationVocal Bursts Intensity Prediction

Inducing Hierarchical Compositional Model by Sparsifying Generator Network

2019-09-10 · CVPR 2020 6 · Xianglei Xing, Tianfu Wu, Song-Chun Zhu, Ying Nian Wu

This paper proposes to learn hierarchical compositional AND-OR model for interpretable image synthesis by sparsifying the generator network. The proposed method adopts the scene-objects-parts-subparts-primitives hierarch…

Image GenerationImage Reconstructionmodel

Geometric Iterative Retrieval for Neural Audio Codec Resynthesis

2026-08-19 · Leo Schmidt-Traub, Frédéric Berdoz, Luca A. Lanzendörfer, Roger Wattenhofer arxiv

Neural audio codecs based on Residual Vector Quantization (RVQ) have become the dominant discrete representation for token-based general audio generation, yet resynthesizing high-quality audio from coarse codec tokens re…

Audio Generation

Conditional Deep Gaussian Processes: multi-fidelity kernel learning

2020-02-07 · Chi-Ken Lu, Patrick Shafto

Deep Gaussian Processes (DGPs) were proposed as an expressive Bayesian model capable of a mathematically grounded estimation of uncertainty. The expressivity of DPGs results from not only the compositional character but …

Few-Shot LearningGaussian ProcessesInductive Biasregression+2

PosterIQ: A Design Perspective Benchmark for Poster Understanding and Generation

2026-03-25 · Yuheng Feng, Wen Zhang, Haodong Duan, Xingxing Zou arxiv

We present PosterIQ, a design-driven benchmark for poster understanding and generation, annotated across composition structure, typographic hierarchy, and semantic intent. It includes 7,765 image-annotation instances and…