paper-with-me

홈 › Papers

ArtiFade: Learning to Generate High-quality Subject from Blemished Images

2024-09-05 · CVPR 2025 1 · Shuya Yang, Shaozhe Hao, Yukang Cao, Kwan-Yee K. Wong

Subject-driven text-to-image generation has witnessed remarkable advancements in its ability to learn and capture characteristics of a subject using only a limited number of images. However, existing methods commonly rely on high-quality images for training and may struggle to generate reasonable images when the input images are blemished by artifacts. This is primarily attributed to the inadequate capability of current techniques in distinguishing subject-related features from disruptive artifacts. In this paper, we introduce ArtiFade to tackle this issue and successfully generate high-quality artifact-free images from blemished datasets. Specifically, ArtiFade exploits fine-tuning of a pre-trained text-to-image model, aiming to remove artifacts. The elimination of artifacts is achieved by utilizing a specialized dataset that encompasses both unblemished images and their corresponding blemished counterparts during fine-tuning. ArtiFade also ensures the preservation of the original generative capabilities inherent within the diffusion model, thereby enhancing the overall performance of subject-driven methods in generating high-quality and artifact-free images. We further devise evaluation benchmarks tailored for this task. Through extensive qualitative and quantitative experiments, we demonstrate the generalizability of ArtiFade in effective artifact removal under both in-distribution and out-of-distribution scenarios.

📄 PDF Abstract BibTeX arXiv:2409.03745

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationText to Image GenerationText-to-Image Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Yuan: Yielding Unblemished Aesthetics Through A Unified Network for Visual Imperfections Removal in Generated Images

2025-01-15 · Zhenyu Yu, Chee Seng Chan

Generative AI presents transformative potential across various domains, from creative arts to scientific visualization. However, the utility of AI-generated imagery is often compromised by visual flaws, including anatomi…

Image Generation

Maximum Entropy Multi-Task Inverse RL

2020-04-27 · Saurabh Arora, Bikramjit Banerjee, Prashant Doshi

Multi-task IRL allows for the possibility that the expert could be switching between multiple ways of solving the same problem, or interleaving demonstrations of multiple tasks. The learner aims to learn the multiple rew…

Clustering

AGIQA-3K: An Open Database for AI-Generated Image Quality Assessment

2023-06-07 · Chunyi Li, ZiCheng Zhang, HaoNing Wu, Wei Sun 외

With the rapid advancements of the text-to-image generative model, AI-generated images (AGIs) have been widely applied to entertainment, education, social media, etc. However, considering the large quality variance among…

Image Quality Assessment

AIGIQA-20K: A Large Database for AI-Generated Image Quality Assessment

2024-04-04 · Chunyi Li, Tengchuan Kou, Yixuan Gao, Yuqin Cao 외

With the rapid advancements in AI-Generated Content (AIGC), AI-Generated Images (AIGIs) have been widely applied in entertainment, education, and social media. However, due to the significant variance in quality among di…

Image Quality Assessment

EvalTalker: Learning to Evaluate Real-Portrait-Driven Multi-Subject Talking Humans

2025-12-01 · Yingjie Zhou, Xilei Zhu, Siyu Ren, Ziyi Zhao 외 arxiv

Speech-driven Talking Human (TH) generation, commonly known as "Talker," currently faces limitations in multi-subject driving capabilities. Extending this paradigm to "Multi-Talker," capable of animating multiple subject…