Npix2Cpix: A GAN-Based Image-to-Image Translation Network With Retrieval- Classification Integration for Watermark Retrieval From Historical Document Images
The identification and restoration of ancient watermarks have long been a major topic in codicology and history. Classifying historical documents based on watermarks is challenging due to their diversity, noisy samples, multiple representation modes, and minor distinctions between classes and intra-class variations. This paper proposes a modified U-net-based conditional generative adversarial network (GAN) named Npix2Cpix to translate noisy raw historical watermarked images into clean, handwriting-free watermarked images by performing image translation from degraded (noisy) pixels to clean pixels. Using image-to-image translation and adversarial learning, the network creates clutter-free images for watermark restoration and categorization. The generator and discriminator of the proposed GAN are trained using two separate loss functions, each based on the distance between images, to learn the mapping from the input noisy image to the output clean image. After using the proposed GAN to pre-process noisy watermarked images, Siamese-based one-shot learning is employed for watermark classification. Experimental results on a large-scale historical watermark dataset demonstrate that cleaning the noisy watermarked images can help to achieve high one-shot classification accuracy. The qualitative and quantitative evaluation of the retrieved watermarked image highlights the effectiveness of the proposed approach.
Code (0)
등록된 구현이 없습니다.
Tasks
Generative Adversarial NetworkImage-to-Image TranslationOne-Shot LearningRetrievalTranslationSimilar Papers 제목 키워드 기반
ReasonPix2Pix: Instruction Reasoning Dataset for Advanced Image Editing
Instruction-based image editing focuses on equipping a generative model with the capacity to adhere to human-written instructions for editing images. Current approaches typically comprehend explicit and specific instruct…
Retrieval Guided Unsupervised Multi-domain Image-to-Image Translation
Image to image translation aims to learn a mapping that transforms an image from one visual domain to another. Recent works assume that images descriptors can be disentangled into a domain-invariant content representatio…
Image RetrievalImage-to-Image TranslationRetrievalTranslationFashion Image-to-Image Translation for Complementary Item Retrieval
The increasing demand for online fashion retail has boosted research in fashion compatibility modeling and item retrieval, focusing on matching user queries (textual descriptions or reference images) with compatible fash…
Image RetrievalImage-to-Image TranslationRetrievalTranslationMultimodal Neural Machine Translation with Search Engine Based Image Retrieval
Recently, numbers of works shows that the performance of neural machine translation (NMT) can be improved to a certain extent with using visual information. However, most of these conclusions are drawn from the analysis …
DescriptiveImage RetrievalMachine TranslationNMT+3Sketch-based Image Retrieval from Millions of Images under Rotation, Translation and Scale Variations
Proliferation of touch-based devices has made sketch-based image retrieval practical. While many methods exist for sketch-based object detection/image retrieval on small datasets, relatively less work has been done on la…
Image RetrievalObjectobject-detectionObject Detection+3