paper-with-me

홈 › Papers

HP-GAN: Harnessing pretrained networks for GAN improvement with FakeTwins and discriminator consistency

2026-02-03 · Geonhui Son, Jeong Ryong Lee, Dosik Hwang arxiv

Generative Adversarial Networks (GANs) have made significant progress in enhancing the quality of image synthesis. Recent methods frequently leverage pretrained networks to calculate perceptual losses or utilize pretrained feature spaces. In this paper, we extend the capabilities of pretrained networks by incorporating innovative self-supervised learning techniques and enforcing consistency between discriminators during GAN training. Our proposed method, named HP-GAN, effectively exploits neural network priors through two primary strategies: FakeTwins and discriminator consistency. FakeTwins leverages pretrained networks as encoders to compute a self-supervised loss and applies this through the generated images to train the generator, thereby enabling the generation of more diverse and high quality images. Additionally, we introduce a consistency mechanism between discriminators that evaluate feature maps extracted from Convolutional Neural Network (CNN) and Vision Transformer (ViT) feature networks. Discriminator consistency promotes coherent learning among discriminators and enhances training robustness by aligning their assessments of image quality. Our extensive evaluation across seventeen datasets-including scenarios with large, small, and limited data, and covering a variety of image domains-demonstrates that HP-GAN consistently outperforms current state-of-the-art methods in terms of Fréchet Inception Distance (FID), achieving significant improvements in image diversity and quality. Code is available at: https://github.com/higun2/HP-GAN.

📄 PDF Abstract BibTeX arXiv:2602.03039

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised Learning

Similar Papers 제목 키워드 기반

SCP-GAN: Self-Correcting Discriminator Optimization for Training Consistency Preserving Metric GAN on Speech Enhancement Tasks

2022-10-26 · Vasily Zadorozhnyy, Qiang Ye, Kazuhito Koishida

In recent years, Generative Adversarial Networks (GANs) have produced significantly improved results in speech enhancement (SE) tasks. They are difficult to train, however. In this work, we introduce several improvements…

Speech Enhancement

Generating time-consistent dynamics with discriminator-guided image diffusion models

2025-05-14 · Philipp Hess, Maximilian Gelbrecht, Christof Schötz, Michael Aich 외

Realistic temporal dynamics are crucial for many video generation, processing and modelling applications, e.g. in computational fluid dynamics, weather prediction, or long-term climate simulations. Video diffusion models…

Video Generation

Discriminator Guidance for Autoregressive Diffusion Models

2023-10-24 · Filip Ekström Kelvinius, Fredrik Lindsten

We introduce discriminator guidance in the setting of Autoregressive Diffusion Models. The use of a discriminator to guide a diffusion process has previously been used for continuous diffusion models, and in this work we…

Consistent Multimodal Generation via A Unified GAN Framework

2023-07-04 · Zhen Zhu, Yijun Li, Weijie Lyu, Krishna Kumar Singh 외

We investigate how to generate multimodal image outputs, such as RGB, depth, and surface normals, with a single generative model. The challenge is to produce outputs that are realistic, and also consistent with each othe…

multimodal generation

Rethinking Style Transformer by Energy-based Interpretation: Adversarial Unsupervised Style Transfer using Pretrained Model

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Style control, content preservation, and fluency determine the quality of text style transfer models. To train on a nonparallel corpus, several existing approaches aim to deceive the style discriminator with an adversari…

Language ModelingLanguage ModellingStyle TransferText Style Transfer