paper-with-me

홈 › Papers

SynArtifact: Classifying and Alleviating Artifacts in Synthetic Images via Vision-Language Model

2024-02-28 · Bin Cao, Jianhao Yuan, Yexin Liu, Jian Li, Shuyang Sun, Jing Liu, Bo Zhao

In the rapidly evolving area of image synthesis, a serious challenge is the presence of complex artifacts that compromise perceptual realism of synthetic images. To alleviate artifacts and improve quality of synthetic images, we fine-tune Vision-Language Model (VLM) as artifact classifier to automatically identify and classify a wide range of artifacts and provide supervision for further optimizing generative models. Specifically, we develop a comprehensive artifact taxonomy and construct a dataset of synthetic images with artifact annotations for fine-tuning VLM, named SynArtifact-1K. The fine-tuned VLM exhibits superior ability of identifying artifacts and outperforms the baseline by 25.66%. To our knowledge, this is the first time such end-to-end artifact classification task and solution have been proposed. Finally, we leverage the output of VLM as feedback to refine the generative model for alleviating artifacts. Visualization results and user study demonstrate that the quality of images synthesized by the refined diffusion model has been obviously improved.

📄 PDF Abstract BibTeX arXiv:2402.18068

Code (1)

BBBiiinnn/SynArtifact 공식 구현

Tasks

Image GenerationLanguage ModelingLanguage Modelling

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Object Classification in Images of Neoclassical Artifacts Using Deep Learning

2017-10-13 · Bernhard Bermeitinger, Maria Christoforaki, Simon Donig, Siegfried Handschuh

In this paper, we report on our efforts for using Deep Learning for classifying artifacts and their features in digital visuals as a part of the Neoclassica framework. It was conceived to provide scholars with new method…

ClassificationDeep LearningGeneral Classification

Proportional Sensitivity in Generative Adversarial Network (GAN)-Augmented Brain Tumor Classification Using Convolutional Neural Network

2025-06-20 · Mahin Montasir Afif, Abdullah Al Noman, K. M. Tahsin Kabir, Md. Mortuza Ahmmed 외

Generative Adversarial Networks (GAN) have shown potential in expanding limited medical imaging datasets. This study explores how different ratios of GAN-generated and real brain tumor MRI images impact the performance o…

Brain Tumor ClassificationGenerative Adversarial NetworkSensitivity

ECG-Image-Kit: A Synthetic Image Generation Toolbox to Facilitate Deep Learning-Based Electrocardiogram Digitization

2023-07-04 · Kshama Kodthalu Shivashankara, Deepanshi, Afagh Mehri Shervedani, Gari D. Clifford 외

Cardiovascular diseases are a major cause of mortality globally, and electrocardiograms (ECGs) are crucial for diagnosing them. Traditionally, ECGs are printed on paper. However, these printouts, even when scanned, are i…

Data AugmentationDecision MakingDenoisingImage Generation+2

Deep Exposure Fusion with Deghosting via Homography Estimation and Attention Learning

2020-04-20 · Sheng-Yeh Chen, Yung-Yu Chuang

Modern cameras have limited dynamic ranges and often produce images with saturated or dark regions using a single exposure. Although the problem could be addressed by taking multiple images with different exposures, expo…

Homography Estimation

Provenance Analysis of Archaeological Artifacts via Multimodal RAG Systems

2025-09-25 · Tuo Zhang, Yuechun Sun, Ruiliang Liu arxiv

In this work, we present a retrieval-augmented generation (RAG)-based system for provenance analysis of archaeological artifacts, designed to support expert reasoning by integrating multimodal retrieval and large vision-…

Semantic Retrieval