paper-with-me

Papers

Compositional diversity in visual concept learning

2023-05-30 · Yanli Zhou, Reuben Feinman, Brenden M. Lake

Humans leverage compositionality to efficiently learn new concepts, understanding how familiar parts can combine together to form novel objects. In contrast, popular computer vision models struggle to make the same types of inferences, requiring more data and generalizing less flexibly than people do. Here, we study these distinctively human abilities across a range of different types of visual composition, examining how people classify and generate ``alien figures'' with rich relational structure. We also develop a Bayesian program induction model which searches for the best programs for generating the candidate visual figures, utilizing a large program space containing different compositional mechanisms and abstractions. In few shot classification tasks, we find that people and the program induction model can make a range of meaningful compositional generalizations, with the model providing a strong account of the experimental data as well as interpretable parameters that reveal human assumptions about the factors invariant to category membership (here, to rotation and changing part attachment). In few shot generation tasks, both people and the models are able to construct compelling novel examples, with people behaving in additional structured ways beyond the model capabilities, e.g. making choices that complete a set or reconfiguring existing parts in highly novel ways. To capture these additional behavioral patterns, we develop an alternative model based on neuro-symbolic program induction: this model also composes new concepts from existing parts yet, distinctively, it utilizes neural network modules to successfully capture residual statistical structure. Together, our behavioral and computational findings show how people and models can produce a rich variety of compositional behavior when classifying and generating visual objects.

📄 PDF Abstract BibTeX arXiv:2305.19374

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityProgram induction

Similar Papers 제목 키워드 기반

ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty

2024-08-26 · Xindi Wu, Dingli Yu, Yangsibo Huang, Olga Russakovsky 외

Compositionality is a critical capability in Text-to-Image (T2I) models, as it reflects their ability to understand and combine multiple concepts from text descriptions. Existing evaluations of compositional capability r…

DiversityImage Generation

Does Data Scaling Lead to Visual Compositional Generalization?

2025-07-09 · Arnas Uselis, Andrea Dittadi, Seong Joon Oh arxiv

Compositional understanding is crucial for human intelligence, yet it remains unclear whether contemporary vision models exhibit it. The dominant machine learning paradigm is built on the premise that scaling data and mo…

Are Object-Centric Representations Better At Compositional Generalization?

2026-02-18 · Ferdinand Kapl, Amir Mohammad Karimi Mamaghan, Maximilian Seitzer, Karl Henrik Johansson 외 arxiv

Compositional generalization, the ability to reason about novel combinations of familiar concepts, is fundamental to human cognition and a critical challenge for machine learning. Object-centric (OC) representations, whi…

Visual Question Answering

SCAN: Learning Hierarchical Compositional Visual Concepts

2017-07-11 · ICLR 2018 1 · Irina Higgins, Nicolas Sonnerat, Loic Matthey, Arka Pal 외

The seemingly infinite diversity of the natural world arises from a relatively small set of coherent rules, such as the laws of physics or chemistry. We conjecture that these rules give rise to regularities that can be d…

MetaReVision: Meta-Learning with Retrieval for Visually Grounded Compositional Concept Acquisition

2023-11-02 · Guangyue Xu, Parisa Kordjamshidi, Joyce Chai

Humans have the ability to learn novel compositional concepts by recalling and generalizing primitive concepts acquired from past experiences. Inspired by this observation, in this paper, we propose MetaReVision, a retri…

Meta-LearningRetrieval