paper-with-me

Papers

Compositional Scene Representation Learning via Reconstruction: A Survey

2022-02-15 · Jinyang Yuan, Tonglin Chen, Bin Li, xiangyang xue

Visual scenes are composed of visual concepts and have the property of combinatorial explosion. An important reason for humans to efficiently learn from diverse visual scenes is the ability of compositional perception, and it is desirable for artificial intelligence to have similar abilities. Compositional scene representation learning is a task that enables such abilities. In recent years, various methods have been proposed to apply deep neural networks, which have been proven to be advantageous in representation learning, to learn compositional scene representations via reconstruction, advancing this research direction into the deep learning era. Learning via reconstruction is advantageous because it may utilize massive unlabeled data and avoid costly and laborious data annotation. In this survey, we first outline the current progress on reconstruction-based compositional scene representation learning with deep neural networks, including development history and categorizations of existing methods from the perspectives of the modeling of visual scenes and the inference of scene representations; then provide benchmarks, including an open source toolbox to reproduce the benchmark experiments, of representative methods that consider the most extensively studied problem setting and form the foundation for other methods; and finally discuss the limitations of existing methods and future directions of this research topic.

📄 PDF Abstract BibTeX arXiv:2202.07135

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningSurvey

Similar Papers 제목 키워드 기반

Object-Compositional Neural Implicit Surfaces

2022-07-20 · Qianyi Wu, Xian Liu, Yuedong Chen, Kejie Li 외

The neural implicit representation has shown its effectiveness in novel view synthesis and high-quality 3D reconstruction from multi-view images. However, most approaches focus on holistic scene representation yet ignore…

3D ReconstructionNovel View SynthesisObject

SimRecon: SimReady Compositional Scene Reconstruction from Real Videos

2026-03-02 · Chong Xia, Kai Zhu, Zizhuo Wang, Fangfu Liu 외 arxiv

Compositional scene reconstruction seeks to create object-centric representations rather than holistic scenes from real-world videos, which is natively applicable for simulation and interaction. Conventional compositiona…

Gaussian Object Carver: Object-Compositional Gaussian Splatting with surfaces completion

2024-12-03 · Liu Liu, Xinjie Wang, Jiaxiong Qiu, Tianwei Lin 외

3D scene reconstruction is a foundational problem in computer vision. Despite recent advancements in Neural Implicit Representations (NIR), existing methods often lack editability and compositional flexibility, limiting …

3D Scene ReconstructionObject

GaussianBlock: Building Part-Aware Compositional and Editable 3D Scene by Primitives and Gaussians

2024-10-02 · Shuyi Jiang, QiHao Zhao, Hossein Rahmani, De Wen Soh 외

Recently, with the development of Neural Radiance Fields and Gaussian Splatting, 3D reconstruction techniques have achieved remarkably high fidelity. However, the latent representations learnt by these methods are highly…

3D Reconstruction

ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment

2026-04-12 · Mingyu Dong, Chong Xia, Mingyuan Jia, Weichen Lyu 외 arxiv

Humans exhibit an innate capacity to rapidly perceive and segment objects from video observations, and even mentally assemble them into structured 3D scenes. Replicating such capability, termed compositional 3D reconstru…

3D Reconstruction