paper-with-me

Papers

Provably Learning Object-Centric Representations

2023-05-23 · Jack Brady, Roland S. Zimmermann, Yash Sharma, Bernhard Schölkopf, Julius von Kügelgen, Wieland Brendel

Learning structured representations of the visual world in terms of objects promises to significantly improve the generalization abilities of current machine learning models. While recent efforts to this end have shown promising empirical progress, a theoretical account of when unsupervised object-centric representation learning is possible is still lacking. Consequently, understanding the reasons for the success of existing object-centric methods as well as designing new theoretically grounded methods remains challenging. In the present work, we analyze when object-centric representations can provably be learned without supervision. To this end, we first introduce two assumptions on the generative process for scenes comprised of several objects, which we call compositionality and irreducibility. Under this generative process, we prove that the ground-truth object representations can be identified by an invertible and compositional inference model, even in the presence of dependencies between objects. We empirically validate our results through experiments on synthetic data. Finally, we provide evidence that our theory holds predictive power for existing object-centric models by showing a close correspondence between models' compositionality and invertibility and their empirical identifiability.

📄 PDF Abstract BibTeX arXiv:2305.14229

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectRepresentation Learning

Similar Papers 제목 키워드 기반

Provable Compositional Generalization for Object-Centric Learning

2023-10-09 · Thaddäus Wiedemer, Jack Brady, Alexander Panfilov, Attila Juhos 외

Learning representations that generalize to novel compositions of known concepts is crucial for bridging the gap between human and machine perception. One prominent effort is learning object-centric representations, whic…

DecoderObject

Learning Global Object-Centric Representations via Disentangled Slot Attention

2024-10-24 · Tonglin Chen, Yinxuan Huang, Zhimeng Shen, Jinghao Huang 외

Humans can discern scene-independent features of objects across various environments, allowing them to swiftly identify objects amidst changing factors such as lighting, perspective, size, and position and imagine the co…

ObjectPositionRepresentation LearningScene Generation

A Data-Centric Revisit of Pre-Trained Vision Models for Robot Learning

2025-03-10 · CVPR 2025 1 · Xin Wen, Bingchen Zhao, Yilun Chen, Jiangmiao Pang 외

Pre-trained vision models (PVMs) are fundamental to modern robotics, yet their optimal configuration remains unclear. Through systematic evaluation, we find that while DINO and iBOT outperform MAE across visuomotor contr…

ObjectScene Understanding

Dyn-O: Building Structured World Models with Object-Centric Representations

2025-07-04 · Zizhao Wang, Kaixin Wang, Li Zhao, Peter Stone 외

World models aim to capture the dynamics of the environment, enabling agents to predict and plan for future states. In most scenarios of interest, the dynamics are highly centered on interactions among objects within the…

Object

CarFormer: Self-Driving with Learned Object-Centric Representations

2024-07-22 · Shadi Hamdan, Fatma Güney

The choice of representation plays a key role in self-driving. Bird's eye view (BEV) representations have shown remarkable performance in recent years. In this paper, we propose to learn object-centric representations in…

Object