FineTag: Multi-attribute Classification at Fine-grained Level in Images
In this paper, we address the extraction of the fine-grained attributes of an instance as a `multi-attribute classification' problem. To this end, we propose an end-to-end architecture by adopting the bi-linear Convolutional Neural Network with the pairwise ranking loss. This is the first time such architecture is applied for the fine-grained attributes classification problem. We compared the proposed method with a competitive deep Convolutional Neural Network baseline. Extensive experiments show that the proposed method attains/outperforms the performance of compared baseline with significantly less number of parameters ($40\times$ less). We demonstrated our approach on CUB200 birds dataset whose annotations are adapted in this work for multi-attribute classification at fine-grained level.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeClassificationGeneral ClassificationSimilar Papers 제목 키워드 기반
Enhancing Fine-Grained Classification for Low Resolution Images
Low resolution fine-grained classification has widespread applicability for applications where data is captured at a distance such as surveillance and mobile photography. While fine-grained classification with high resol…
AttributeClassificationGeneral ClassificationHierarchical Fine-Grained Image Forgery Detection and Localization
Differences in forgery attributes of images generated in CNN-synthesized and image-editing domains are large, and such differences make a unified image forgery detection and localization (IFDL) challenging. To this end, …
AttributeClassificationImage Forgery DetectionRepresentation LearningLanguage-guided Hierarchical Fine-grained Image Forgery Detection and Localization
Differences in forgery attributes of images generated in CNN-synthesized and image-editing domains are large, and such differences make a unified image forgery detection and localization (IFDL) challenging. To this end, …
AttributeImage Forgery DetectionRepresentation LearningZero-Shot Product Attribute Labeling with Vision-Language Models: A Three-Tier Evaluation Framework
Fine-grained attribute prediction is essential for fashion retail applications including catalog enrichment, visual search, and recommendation systems. Vision-Language Models (VLMs) offer zero-shot prediction without tas…
Recommendation SystemsReal Classification by Description: Extending CLIP's Limits of Part Attributes Recognition
In this study, we define and tackle zero shot "real" classification by description, a novel task that evaluates the ability of Vision-Language Models (VLMs) like CLIP to classify objects based solely on descriptive attri…
AttributeDescriptiveObjectObject Recognition+1