Weakly Supervised Learning of Objects, Attributes and their Associations
When humans describe images they tend to use combinations of nouns and adjectives, corresponding to objects and their associated attributes respectively. To generate such a description automatically, one needs to model objects, attributes and their associations. Conventional methods require strong annotation of object and attribute locations, making them less scalable. In this paper, we model object-attribute associations from weakly labelled images, such as those widely available on media sharing sites (e.g. Flickr), where only image-level labels (either object or attributes) are given, without their locations and associations. This is achieved by introducing a novel weakly supervised non-parametric Bayesian model. Once learned, given a new image, our model can describe the image, including objects, attributes and their associations, as well as their locations and segmentation. Extensive experiments on benchmark datasets demonstrate that our weakly supervised model performs at par with strongly supervised models on tasks such as image description and retrieval based on object-attribute associations.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeImage DescriptionObjectRetrievalWeakly-supervised LearningSimilar Papers 제목 키워드 기반
Weakly Supervised Image Annotation and Segmentation with Objects and Attributes
We propose to model complex visual scenes using a non-parametric Bayesian model learned from weakly labelled images abundant on media sharing sites such as Flickr. Given weak image-level annotations of objects and attrib…
AttributeObjectobject-detectionObject Detection+3LOCL: Learning Object-Attribute Composition using Localization
This paper describes LOCL (Learning Object Attribute Composition using Localization) that generalizes composition zero shot learning to objects in cluttered and more realistic settings. The problem of unseen Object Attri…
AttributeObjectZero-Shot LearningWeakly-Supervised End-to-End CAD Retrieval to Scan Objects
CAD model retrieval to real-world scene observations has shown strong promise as a basis for 3D perception of objects and a clean, lightweight mesh-based scene representation; however, current approaches to retrieve CAD …
object-detectionObject DetectionRetrievalRecovering the Missing Link: Predicting Class-Attribute Associations for Unsupervised Zero-Shot Learning
Collecting training images for all visual categories is not only expensive but also impractical. Zero-shot learning (ZSL), especially using attributes, offers a pragmatic solution to this problem. However, at test time m…
Attributezero-shot-classificationZero-Shot LearningWeakly Supervised Learning for Attribute Localization in Outdoor Scenes
In this paper, we propose a weakly supervised method for simultaneously learning scene parts and attributes from a collection of images associated with attributes in text, where the precise localization of the each attri…
AttributeWeakly-supervised Learning