Toward Spatially Unbiased Generative Models
Recent image generation models show remarkable generation performance. However, they mirror strong location preference in datasets, which we call spatial bias. Therefore, generators render poor samples at unseen locations and scales. We argue that the generators rely on their implicit positional encoding to render spatial content. From our observations, the generator's implicit positional encoding is translation-variant, making the generator spatially biased. To address this issue, we propose injecting explicit positional encoding at each scale of the generator. By learning the spatially unbiased generator, we facilitate the robust use of generators in multiple tasks, such as GAN inversion, multi-scale generation, generation of arbitrary sizes and aspect ratios. Furthermore, we show that our method can also be applied to denoising diffusion probabilistic models.
Code (2)
Tasks
DenoisingImage GenerationTranslationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Automated data-driven selection of the hyperparameters for Total-Variation based texture segmentation
Penalized Least Squares are widely used in signal and image processing. Yet, it suffers from a major limitation since it requires fine-tuning of the regularization parameters. Under assumptions on the noise probability d…
Transcriptome-supervised classification of tissue morphology using deep learning
Deep learning has proven to successfully learn variations in tissue and cell morphology. Training of such models typically relies on expensive manual annotations. Here we conjecture that spatially resolved gene expressio…
Deep LearningStatistically unbiased prediction enables accurate denoising of voltage imaging data
Here we report SUPPORT (Statistically Unbiased Prediction utilizing sPatiOtempoRal information in imaging daTa), a self-supervised learning method for removing Poisson-Gaussian noise in voltage imaging data. SUPPORT is b…
DenoisingSelf-Supervised LearningAsymptotically unbiased estimation of physical observables with neural samplers
We propose a general framework for the estimation of observables with generative neural samplers focusing on modern deep generative neural networks that provide an exact sampling probability. In this framework, we presen…
Relative Transfer Function Estimation Exploiting Spatially Separated Microphones in a Diffuse Noise Field
Many multi-microphone speech enhancement algorithms require the relative transfer function (RTF) vector of the desired speech source, relating the acoustic transfer functions of all array microphones to a reference micro…
Speech Enhancement