Beyond Seen Primitive Concepts and Attribute-Object Compositional Learning
Learning from seen attribute-object pairs to generalize to unseen compositions has been studied extensively in Compositional Zero-Shot Learning (CZSL). However CZSL setup is still limited to seen attributes and objects and cannot generalize to unseen concepts and their compositions. To overcome this limitation we propose a new task Open Vocabulary-Compositional Zero-shot Learning (OV-CZSL) where unseen attributes objects and unseen compositions are evaluated. To show that OV-CZSL is a challenging yet solvable problem we propose three new benchmarks based on existing datasets MIT-States C-GQA and VAW-CZSL along with new baselines and evaluation setup. We use language embeddings and external vocabulary with our novel neighborhood expansion loss to allow any method to learn semantic correlations between seen and unseen primitives.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeCompositional Zero-Shot LearningZero-Shot LearningSimilar Papers 제목 키워드 기반
Relation-aware Compositional Zero-shot Learning for Attribute-Object Pair Recognition
This paper proposes a novel model for recognizing images with composite attribute-object concepts, notably for composite concepts that are unseen during model training. We aim to explore the three key properties required…
AttributeBlockingCompositional Zero-Shot LearningRelation+1Learning Clustering-based Prototypes for Compositional Zero-shot Learning
Learning primitive (i.e., attribute and object) concepts from seen compositions is the primary challenge of Compositional Zero-Shot Learning (CZSL). Existing CZSL solutions typically rely on oversimplified data assumptio…
AttributeClusteringCompositional Zero-Shot LearningContrastive Learning+1LOGICZSL: Exploring Logic-induced Representation for Compositional Zero-shot Learning
Compositional zero-shot learning (CZSL) aims to recognize unseen attribute-object compositions by learning the primitive concepts (*i.e.*, attribute and object) from the training set. While recent works achieve impre…
AttributeCompositional Zero-Shot LearningZero-Shot LearningAttention Based Simple Primitives for Open World Compositional Zero-Shot Learning
Compositional Zero-Shot Learning (CZSL) aims to predict unknown compositions made up of attribute and object pairs. Predicting compositions unseen during training is a challenging task. We are exploring Open World Compos…
AttributeCompositional Zero-Shot LearningObjectZero-Shot LearningDo Vision-Language Pretrained Models Learn Composable Primitive Concepts?
Vision-language (VL) pretrained models have achieved impressive performance on multimodal reasoning and zero-shot recognition tasks. Many of these VL models are pretrained on unlabeled image and caption pairs from the in…
Fine-Grained Visual RecognitionMultimodal ReasoningZero-Shot Learning