paper-with-me

홈 › Papers

Learning Compositional Visual Concepts with Mutual Consistency

2017-11-16 · CVPR 2018 6 · Yunye Gong, Srikrishna Karanam, Ziyan Wu, Kuan-Chuan Peng, Jan Ernst, Peter C. Doerschuk

Compositionality of semantic concepts in image synthesis and analysis is appealing as it can help in decomposing known and generatively recomposing unknown data. For instance, we may learn concepts of changing illumination, geometry or albedo of a scene, and try to recombine them to generate physically meaningful, but unseen data for training and testing. In practice however we often do not have samples from the joint concept space available: We may have data on illumination change in one data set and on geometric change in another one without complete overlap. We pose the following question: How can we learn two or more concepts jointly from different data sets with mutual consistency where we do not have samples from the full joint space? We present a novel answer in this paper based on cyclic consistency over multiple concepts, represented individually by generative adversarial networks (GANs). Our method, ConceptGAN, can be understood as a drop in for data augmentation to improve resilience for real world applications. Qualitative and quantitative evaluations demonstrate its efficacy in generating semantically meaningful images, as well as one shot face verification as an example application.

📄 PDF Abstract BibTeX arXiv:1711.06148

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationFace VerificationImage Generation

Similar Papers 제목 키워드 기반

Maintaining Reasoning Consistency in Compositional Visual Question Answering

2022-01-01 · CVPR 2022 1 · Chenchen Jing, Yunde Jia, Yuwei Wu, Xinyu Liu 외

A compositional question refers to a question that contains multiple visual concepts (e.g., objects, attributes, and relationships) and requires compositional reasoning to answer. Existing VQA models can answer a com…

Question AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Gate-and-Merge: Zero-shot Compositional Personalization of Vision Language Models

2026-05-09 · Guodong Ding, Angela Yao arxiv

This paper tackles compositional personalization of vision-language models (VLMs). In this problem, multiple user-defined concepts must be recognized or described jointly at test time. We introduce Gate-and-Merge, a zero…

Consistency of Compositional Generalization across Multiple Levels

2024-12-18 · Chuanhao Li, Zhen Li, Chenchen Jing, Xiaomeng Fan 외

Compositional generalization is the capability of a model to understand novel compositions composed of seen concepts. There are multiple levels of novel compositions including phrase-phrase level, phrase-word level, and …

Meta-LearningQuestion AnsweringVideo GroundingVisual Question Answering

Flexible Compositional Learning of Structured Visual Concepts

2021-05-20 · Yanli Zhou, Brenden M. Lake

Humans are highly efficient learners, with the ability to grasp the meaning of a new concept from just a few examples. Unlike popular computer vision systems, humans can flexibly leverage the compositional structure of t…

Program induction

MetaReVision: Meta-Learning with Retrieval for Visually Grounded Compositional Concept Acquisition

2023-11-02 · Guangyue Xu, Parisa Kordjamshidi, Joyce Chai

Humans have the ability to learn novel compositional concepts by recalling and generalizing primitive concepts acquired from past experiences. Inspired by this observation, in this paper, we propose MetaReVision, a retri…

Meta-LearningRetrieval