paper-with-me

홈 › Papers

OC-NMN: Object-centric Compositional Neural Module Network for Generative Visual Analogical Reasoning

2023-10-28 · Rim Assouel, Pau Rodriguez, Perouz Taslakian, David Vazquez, Yoshua Bengio

A key aspect of human intelligence is the ability to imagine -- composing learned concepts in novel ways -- to make sense of new scenarios. Such capacity is not yet attained for machine learning systems. In this work, in the context of visual reasoning, we show how modularity can be leveraged to derive a compositional data augmentation framework inspired by imagination. Our method, denoted Object-centric Compositional Neural Module Network (OC-NMN), decomposes visual generative reasoning tasks into a series of primitives applied to objects without using a domain-specific language. We show that our modular architectural choices can be used to generate new training tasks that lead to better out-of-distribution generalization. We compare our model to existing and new baselines in proposed visual reasoning benchmark that consists of applying arithmetic operations to MNIST digits.

📄 PDF Abstract BibTeX arXiv:2310.18807

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationOut-of-Distribution GeneralizationVisual Reasoning

Similar Papers 제목 키워드 기반

Robust and Controllable Object-Centric Learning through Energy-based Models

2022-10-11 · Ruixiang Zhang, Tong Che, Boris Ivanovic, Renhao Wang 외

Humans are remarkably good at understanding and reasoning about complex visual scenes. The capability to decompose low-level observations into discrete objects allows us to build a grounded abstract representation and id…

ObjectRepresentation LearningScene Generation

Provably Learning Object-Centric Representations

2023-05-23 · Jack Brady, Roland S. Zimmermann, Yash Sharma, Bernhard Schölkopf 외

Learning structured representations of the visual world in terms of objects promises to significantly improve the generalization abilities of current machine learning models. While recent efforts to this end have shown p…

ObjectRepresentation Learning

Slot-VAE: Object-Centric Scene Generation with Slot Attention

2023-06-12 · Yanbo Wang, Letao Liu, Justin Dauwels

Slot attention has shown remarkable object-centric representation learning performance in computer vision tasks without requiring any supervision. Despite its object-centric binding ability brought by compositional model…

ObjectRepresentation LearningScene Generation

GENESIS: Generative Scene Inference and Sampling with Object-Centric Latent Representations

2019-07-30 · ICLR 2020 1 · Martin Engelcke, Adam R. Kosiorek, Oiwi Parker Jones, Ingmar Posner

Generative latent-variable models are emerging as promising tools in robotics and reinforcement learning. Yet, even though tasks in these domains typically involve distinct objects, most state-of-the-art generative model…

Image GenerationObject DiscoveryReinforcement LearningRepresentation Learning+3

Compositional Video Synthesis by Temporal Object-Centric Learning

2025-07-28 · Adil Kaan Akan, Yucel Yemez arxiv

We present a novel framework for compositional video synthesis that leverages temporally consistent object-centric representations, extending our previous work, SlotAdapt, from images to video. While existing object-cent…

Scene UnderstandingVideo Generation