On the Relationship Between Visual Attributes and Convolutional Networks
One of the cornerstone principles of deep models is their abstraction capacity, i.e. their ability to learn abstract concepts from `simpler' ones. Through extensive experiments, we characterize the nature of the relationship between abstract concepts (specifically objects in images) learned by popular and high performing convolutional networks (conv-nets) and established mid-level representations used in computer vision (specifically semantic visual attributes). We focus on attributes due to their impact on several applications, such as object description, retrieval and mining, and active (and zero-shot) learning. Among the findings we uncover, we show empirical evidence of the existence of Attribute Centric Nodes (ACNs) within a conv-net, which is trained to recognize objects (not attributes) in images. These special conv-net nodes (1) collectively encode information pertinent to visual attribute representation and discrimination, (2) are unevenly and sparsely distribution across all layers of the conv-net, and (3) play an important role in conv-net based object recognition.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeObject RecognitionRetrievalZero-Shot LearningSimilar Papers 제목 키워드 기반
Dual Relation Mining Network for Zero-Shot Learning
Zero-shot learning (ZSL) aims to recognize novel classes through transferring shared semantic knowledge (e.g., attributes) from seen classes to unseen classes. Recently, attention-based methods have exhibited significant…
AttributeRelationTransfer LearningZero-Shot LearningVisual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations
Despite progress in perceptual tasks such as image classification, computers still perform poorly on cognitive tasks such as image description and question answering. Cognition is core to tasks that involve not just reco…
image-classificationImage ClassificationImage DescriptionQuestion AnsweringPedestrian Attribute Recognition via Hierarchical Cross-Modality HyperGraph Learning
Current Pedestrian Attribute Recognition (PAR) algorithms typically focus on mapping visual features to semantic labels or attempt to enhance learning by fusing visual and attribute information. However, these methods fa…
Pedestrian Attribute RecognitionGraph-based Visual-Semantic Entanglement Network for Zero-shot Image Recognition
Zero-shot learning uses semantic attributes to connect the search space of unseen objects. In recent years, although the deep convolutional network brings powerful visual modeling capabilities to the ZSL task, its visual…
AttributeZero-Shot LearningContext-Aware Embeddings for Automatic Art Analysis
Automatic art analysis aims to classify and retrieve artistic representations from a collection of images by using computer vision and machine learning techniques. In this work, we propose to enhance visual representatio…
Art AnalysisCross-Modal RetrievalGeneral ClassificationMulti-Task Learning+1