paper-with-me

홈 › Papers

Patch-enhanced Mask Encoder Prompt Image Generation

2024-05-29 · Shusong Xu, Peiye Liu

Artificial Intelligence Generated Content(AIGC), known for its superior visual results, represents a promising mitigation method for high-cost advertising applications. Numerous approaches have been developed to manipulate generated content under different conditions. However, a crucial limitation lies in the accurate description of products in advertising applications. Applying previous methods directly may lead to considerable distortion and deformation of advertised products, primarily due to oversimplified content control conditions. Hence, in this work, we propose a patch-enhanced mask encoder approach to ensure accurate product descriptions while preserving diverse backgrounds. Our approach consists of three components Patch Flexible Visibility, Mask Encoder Prompt Adapter and an image Foundation Model. Patch Flexible Visibility is used for generating a more reasonable background image. Mask Encoder Prompt Adapter enables region-controlled fusion. We also conduct an analysis of the structure and operational mechanisms of the Generation Module. Experimental results show our method can achieve the highest visual results and FID scores compared with other methods.

📄 PDF Abstract BibTeX arXiv:2405.19085

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

Adapter 설명 없음

Similar Papers 제목 키워드 기반

Vision-guided and Mask-enhanced Adaptive Denoising for Prompt-based Image Editing

2024-10-14 · Kejie Wang, Xuemeng Song, Meng Liu, Jin Yuan 외

Text-to-image diffusion models have demonstrated remarkable progress in synthesizing high-quality images from text prompts, which boosts researches on prompt-based image editing that edits a source image according to a t…

DenoisingImage GenerationText-based Image Editing

MAEDiff: Masked Autoencoder-enhanced Diffusion Models for Unsupervised Anomaly Detection in Brain Images

2024-01-19 · Rui Xu, Yunke Wang, Bo Du

Unsupervised anomaly detection has gained significant attention in the field of medical imaging due to its capability of relieving the costly pixel-level annotation. To achieve this, modern approaches usually utilize gen…

Anomaly DetectionUnsupervised Anomaly Detection

3D Masked Autoencoders with Application to Anomaly Detection in Non-Contrast Enhanced Breast MRI

2023-03-10 · Daniel M. Lang, Eli Schwartz, Cosmin I. Bercea, Raja Giryes 외

Self-supervised models allow (pre-)training on unlabeled data and therefore have the potential to overcome the need for large annotated cohorts. One leading self-supervised model is the masked autoencoder (MAE) which was…

Anomaly DetectionLesion Detection

Context Autoencoder for Self-Supervised Representation Learning

2022-02-07 · Xiaokang Chen, Mingyu Ding, Xiaodi Wang, Ying Xin 외

We present a novel masked image modeling (MIM) approach, context autoencoder (CAE), for self-supervised representation pretraining. We pretrain an encoder by making predictions in the encoded representation space. The pr…

DecoderInstance Segmentationobject-detectionObject Detection+4

MAESIL: Masked Autoencoder for Enhanced Self-supervised Medical Image Learning

2026-04-01 · Kyeonghun Kim, Hyeonseok Jung, Youngung Han, Junsu Lim 외 arxiv

Training deep learning models for three-dimensional (3D) medical imaging, such as Computed Tomography (CT), is fundamentally challenged by the scarcity of labeled data. While pre-training on natural images is common, it …

Self-Supervised LearningComputational Efficiency