paper-with-me

홈 › Papers

Generating High Fidelity Images with Subscale Pixel Networks and Multidimensional Upscaling

2018-12-04 · ICLR 2019 5 · Jacob Menick, Nal Kalchbrenner

The unconditional generation of high fidelity images is a longstanding benchmark for testing the performance of image decoders. Autoregressive image models have been able to generate small images unconditionally, but the extension of these methods to large images where fidelity can be more readily assessed has remained an open problem. Among the major challenges are the capacity to encode the vast previous context and the sheer difficulty of learning a distribution that preserves both global semantic coherence and exactness of detail. To address the former challenge, we propose the Subscale Pixel Network (SPN), a conditional decoder architecture that generates an image as a sequence of sub-images of equal size. The SPN compactly captures image-wide spatial dependencies and requires a fraction of the memory and the computation required by other fully autoregressive models. To address the latter challenge, we propose to use Multidimensional Upscaling to grow an image in both size and depth via intermediate stages utilising distinct SPNs. We evaluate SPNs on the unconditional generation of CelebAHQ of size 256 and of ImageNet from size 32 to 256. We achieve state-of-the-art likelihood results in multiple settings, set up new benchmark results in previously unexplored settings and are able to generate very high fidelity large scale samples on the basis of both datasets.

📄 PDF Abstract BibTeX arXiv:1812.01608

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderImage GenerationVocal Bursts Intensity Prediction

Similar Papers 제목 키워드 기반

Pixal3D: Pixel-Aligned 3D Generation from Images

2026-05-11 · Dong-Yang Li, Wang Zhao, Yuxin Chen, Wenbo Hu 외 arxiv

Recent advances in 3D generative models have rapidly improved image-to-3D synthesis quality, enabling higher-resolution geometry and more realistic appearance. Yet fidelity, which measures pixel-level faithfulness of the…

3D Reconstruction3D Generation

HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images

2026-03-02 · Yichen Liu, Donghao Zhou, Jie Wang, Xin Gao 외 arxiv

Human-product images, which showcase the integration of humans and products, play a vital role in advertising, e-commerce, and digital marketing. The essential challenge of generating such images lies in ensuring the hig…

Generating Annotated High-Fidelity Images Containing Multiple Coherent Objects

2020-06-22 · Bryan G. Cardenas, Devanshu Arya, Deepak K. Gupta

Recent developments related to generative models have made it possible to generate diverse high-fidelity images. In particular, layout-to-image generation models have gained significant attention due to their capability …

Image GenerationLayout-to-Image GenerationVocal Bursts Intensity Prediction

PixelRush: Ultra-Fast, Training-Free High-Resolution Image Generation via One-step Diffusion

2026-02-13 · Hong-Phuc Lai, Phong Nguyen, Anh Tran arxiv

Pre-trained diffusion models excel at generating high-quality images but remain inherently limited by their native training resolution. Recent training-free approaches have attempted to overcome this constraint by introd…

Text-to-Image Generation

Training-free Composite Scene Generation for Layout-to-Image Synthesis

2024-07-18 · Jiaqi Liu, Tao Huang, Chang Xu

Recent breakthroughs in text-to-image diffusion models have significantly advanced the generation of high-fidelity, photo-realistic images from textual descriptions. Yet, these models often struggle with interpreting spa…

Image GenerationLayout-to-Image GenerationScene Generation