Fine-grained Recognition in the Wild: A Multi-Task Domain Adaptation Approach
While fine-grained object recognition is an important problem in computer vision, current models are unlikely to accurately classify objects in the wild. These fully supervised models need additional annotated images to classify objects in every new scenario, a task that is infeasible. However, sources such as e-commerce websites and field guides provide annotated images for many classes. In this work, we study fine-grained domain adaptation as a step towards overcoming the dataset shift between easily acquired annotated images and the real world. Adaptation has not been studied in the fine-grained setting where annotations such as attributes could be used to increase performance. Our work uses an attribute based multi-task adaptation loss to increase accuracy from a baseline of 4.1% to 19.1% in the semi-supervised adaptation case. Prior do- main adaptation works have been benchmarked on small datasets such as [46] with a total of 795 images for some domains, or simplistic datasets such as [41] consisting of digits. We perform experiments on a subset of a new challenging fine-grained dataset consisting of 1,095,021 images of 2, 657 car categories drawn from e-commerce web- sites and Google Street View.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeDomain AdaptationObject RecognitionSimilar Papers 제목 키워드 기반
Extremely Fine-Grained Visual Classification over Resembling Glyphs in the Wild
Text recognition in the wild is an important technique for digital maps and urban scene understanding, in which the natural resembling properties between glyphs is one of the major reasons that lead to wrong recognition …
Contrastive LearningFine-Grained Image ClassificationFine-Grained Visual RecognitionScene UnderstandingA Fine-Grained Visual Attention Approach for Fingerspelling Recognition in the Wild
Fingerspelling in sign language has been the means of communicating technical terms and proper nouns when they do not have dedicated sign language gestures. Automatic recognition of fingerspelling can help resolve commun…
Optical Flow EstimationHSEmotion Team at ABAW-10 Competition: Facial Expression Recognition, Valence-Arousal Estimation, Action Unit Detection and Fine-Grained Violence Classification
This article presents our results for the 10th Affective Behavior Analysis in-the-Wild (ABAW) competition. For frame-wise facial emotion understanding tasks (frame-wise facial expression recognition, valence-arousal esti…
Facial Expression RecognitionAction Unit DetectionVideo ClassificationEmotion RecognitionUsing Self-Supervised Auxiliary Tasks to Improve Fine-Grained Facial Representation
In this paper, at first, the impact of ImageNet pre-training on fine-grained Facial Emotion Recognition (FER) is investigated which shows that when enough augmentations on images are applied, training from scratch provid…
Emotion RecognitionFacial Emotion RecognitionFacial Expression Recognition (FER)Head Pose Estimation+3Fine-grained Recognition in the Noisy Wild: Sensitivity Analysis of Convolutional Neural Networks Approaches
In this paper, we study the sensitivity of CNN outputs with respect to image transformations and noise in the area of fine-grained recognition. In particular, we answer the following questions (1) how sensitive are CNNs …
Sensitivity