paper-with-me

Papers

Design Booster: A Text-Guided Diffusion Model for Image Translation with Spatial Layout Preservation

2023-02-05 · Shiqi Sun, Shancheng Fang, Qian He, Wei Liu

Diffusion models are able to generate photorealistic images in arbitrary scenes. However, when applying diffusion models to image translation, there exists a trade-off between maintaining spatial structure and high-quality content. Besides, existing methods are mainly based on test-time optimization or fine-tuning model for each input image, which are extremely time-consuming for practical applications. To address these issues, we propose a new approach for flexible image translation by learning a layout-aware image condition together with a text condition. Specifically, our method co-encodes images and text into a new domain during the training phase. In the inference stage, we can choose images/text or both as the conditions for each time step, which gives users more flexible control over layout and content. Experimental comparisons of our method with state-of-the-art methods demonstrate our model performs best in both style image translation and semantic image translation and took the shortest time.

📄 PDF Abstract BibTeX arXiv:2302.02284

Code (0)

등록된 구현이 없습니다.

Tasks

Translation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion

2026-05-18 · Lichen Ma, Zipeng Guo, Yu He, Xiaolong Fu 외 arxiv

To circumvent the inherent fidelity bottlenecks and optimization misalignment of VAE-based latent diffusion, pixel-space diffusion models have emerged as a compelling end-to-end paradigm. However, existing pixel diffusio…

Computational Efficiency

AdBooster: Personalized Ad Creative Generation using Stable Diffusion Outpainting

2023-09-08 · Veronika Shilova, Ludovic Dos Santos, Flavian vasile, Gaëtan Racic 외

In digital advertising, the selection of the optimal item (recommendation) and its best creative presentation (creative optimization) have traditionally been considered separate disciplines. However, both contribute sign…

Data Augmentation

HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions

2024-03-27 · CVPR 2024 1 · Hao Xu, Haipeng Li, Yinqiao Wang, Shuaicheng Liu 외

Reconstructing 3D hand mesh robustly from a single image is very challenging, due to the lack of diversity in existing real-world datasets. While data synthesis helps relieve the issue, the syn-to-real gap still hinders …

3D Hand Pose EstimationDiversity

FusionBooster: A Unified Image Fusion Boosting Paradigm

2023-05-10 · Chunyang Cheng, Tianyang Xu, Xiao-Jun Wu, Hui Li 외

In recent years, numerous ideas have emerged for designing a mutually reinforcing mechanism or extra stages for the image fusion task, ignoring the inevitable gaps between different vision tasks and the computational bur…

EG-Booster: Explanation-Guided Booster of ML Evasion Attacks

2021-08-31 · Abderrahmen Amich, Birhanu Eshete

The widespread usage of machine learning (ML) in a myriad of domains has raised questions about its trustworthiness in security-critical environments. Part of the quest for trustworthy ML is robustness evaluation of ML m…

image-classificationImage Classification