paper-with-me

홈 › Papers

Exploiting CLIP-based Multi-modal Approach for Artwork Classification and Retrieval

2023-09-21 · Alberto Baldrati, Marco Bertini, Tiberio Uricchio, Alberto del Bimbo

Given the recent advances in multimodal image pretraining where visual models trained with semantically dense textual supervision tend to have better generalization capabilities than those trained using categorical attributes or through unsupervised techniques, in this work we investigate how recent CLIP model can be applied in several tasks in artwork domain. We perform exhaustive experiments on the NoisyArt dataset which is a dataset of artwork images crawled from public resources on the web. On such dataset CLIP achieves impressive results on (zero-shot) classification and promising results in both artwork-to-artwork and description-to-artwork domain.

📄 PDF Abstract BibTeX arXiv:2309.12110

Code (0)

등록된 구현이 없습니다.

Tasks

Retrievalzero-shot-classificationZero-Shot Learning

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

ArtSeek: Deep artwork understanding via multimodal in-context reasoning and late interaction retrieval

2025-07-29 · Nicola Fanelli, Gennaro Vessio, Giovanna Castellano arxiv

Analyzing digitized artworks presents unique challenges, requiring not only visual interpretation but also a deep understanding of rich artistic, contextual, and historical knowledge. We introduce ArtSeek, a multimodal f…

Visual Question AnsweringMultimodal Reasoning

Draw Your Art Dream: Diverse Digital Art Synthesis with Multimodal Guided Diffusion

2022-09-27 · Nisha Huang, Fan Tang, WeiMing Dong, Changsheng Xu

Digital art synthesis is receiving increasing attention in the multimedia community because of engaging the public with art effectively. Current digital art synthesis methods usually use single-modality inputs as guidanc…

Diversity

CLIP-Art: Contrastive Pre-training for Fine-Grained Art Classification

2022-04-29 · Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops 2021 6 · Marcos V. Conde, Kerem Turgutlu

Existing computer vision research in artwork struggles with artwork's fine-grained attributes recognition and lack of curated annotated datasets due to their costly creation. To the best of our knowledge, we are one of t…

AttributeClassificationContrastive LearningFine-Grained Visual Recognition+2

Harnessing Self-Supervised Features for Art Classification

2026-05-18 · Federico Melis, Davide Bilardello, Emanuele Prato, Evelyn Turri 외 arxiv

Classifying artworks presents a significant challenge due to the complex interplay of fine-grained details and abstract features that condition the style or genre of an artwork. This paper presents a systematic investiga…

Have Large Vision-Language Models Mastered Art History?

2024-09-05 · Ombretta Strafforello, Derya Soydaner, Michiel Willems, Anne-Sofie Maerten 외

The emergence of large Vision-Language Models (VLMs) has recently established new baselines in image classification across multiple domains. However, the performance of VLMs in the specific task of artwork classification…

Classificationimage-classificationImage Classificationzero-shot-classification+1