paper-with-me

홈 › Papers

Npix2Cpix: A GAN-Based Image-to-Image Translation Network With Retrieval- Classification Integration for Watermark Retrieval From Historical Document Images

2024-06-05 · Utsab Saha, Sawradip Saha, Shaikh Anowarul Fattah, Mohammad Saquib

The identification and restoration of ancient watermarks have long been a major topic in codicology and history. Classifying historical documents based on watermarks is challenging due to their diversity, noisy samples, multiple representation modes, and minor distinctions between classes and intra-class variations. This paper proposes a modified U-net-based conditional generative adversarial network (GAN) named Npix2Cpix to translate noisy raw historical watermarked images into clean, handwriting-free watermarked images by performing image translation from degraded (noisy) pixels to clean pixels. Using image-to-image translation and adversarial learning, the network creates clutter-free images for watermark restoration and categorization. The generator and discriminator of the proposed GAN are trained using two separate loss functions, each based on the distance between images, to learn the mapping from the input noisy image to the output clean image. After using the proposed GAN to pre-process noisy watermarked images, Siamese-based one-shot learning is employed for watermark classification. Experimental results on a large-scale historical watermark dataset demonstrate that cleaning the noisy watermarked images can help to achieve high one-shot classification accuracy. The qualitative and quantitative evaluation of the retrieved watermarked image highlights the effectiveness of the proposed approach.

📄 PDF Abstract BibTeX arXiv:2406.03556

Code (0)

등록된 구현이 없습니다.

Tasks

Generative Adversarial NetworkImage-to-Image TranslationOne-Shot LearningRetrievalTranslation

Similar Papers 제목 키워드 기반

ReasonPix2Pix: Instruction Reasoning Dataset for Advanced Image Editing

2024-05-18 · Ying Jin, Pengyang Ling, Xiaoyi Dong, Pan Zhang 외

Instruction-based image editing focuses on equipping a generative model with the capacity to adhere to human-written instructions for editing images. Current approaches typically comprehend explicit and specific instruct…

Retrieval Guided Unsupervised Multi-domain Image-to-Image Translation

2020-08-11 · Raul Gomez, Yahui Liu, Marco De Nadai, Dimosthenis Karatzas 외

Image to image translation aims to learn a mapping that transforms an image from one visual domain to another. Recent works assume that images descriptors can be disentangled into a domain-invariant content representatio…

Image RetrievalImage-to-Image TranslationRetrievalTranslation

Fashion Image-to-Image Translation for Complementary Item Retrieval

2024-08-19 · Matteo Attimonelli, Claudio Pomo, Dietmar Jannach, Tommaso Di Noia

The increasing demand for online fashion retail has boosted research in fashion compatibility modeling and item retrieval, focusing on matching user queries (textual descriptions or reference images) with compatible fash…

Image RetrievalImage-to-Image TranslationRetrievalTranslation

Multimodal Neural Machine Translation with Search Engine Based Image Retrieval

2022-07-26 · WAT 2022 10 · Zhenhao Tang, Xiaobing Zhang, Zi Long, Xianghua Fu

Recently, numbers of works shows that the performance of neural machine translation (NMT) can be improved to a certain extent with using visual information. However, most of these conclusions are drawn from the analysis …

DescriptiveImage RetrievalMachine TranslationNMT+3

Sketch-based Image Retrieval from Millions of Images under Rotation, Translation and Scale Variations

2015-10-31 · Sarthak Parui, Anurag Mittal

Proliferation of touch-based devices has made sketch-based image retrieval practical. While many methods exist for sketch-based object detection/image retrieval on small datasets, relatively less work has been done on la…

Image RetrievalObjectobject-detectionObject Detection+3