paper-with-me

Papers

Weakly Supervised Learning of Multi-Object 3D Scene Decompositions Using Deep Shape Priors

2020-10-08 · Cathrin Elich, Martin R. Oswald, Marc Pollefeys, Joerg Stueckler

Representing scenes at the granularity of objects is a prerequisite for scene understanding and decision making. We propose PriSMONet, a novel approach based on Prior Shape knowledge for learning Multi-Object 3D scene decomposition and representations from single images. Our approach learns to decompose images of synthetic scenes with multiple objects on a planar surface into its constituent scene objects and to infer their 3D properties from a single view. A recurrent encoder regresses a latent representation of 3D shape, pose and texture of each object from an input RGB image. By differentiable rendering, we train our model to decompose scenes from RGB-D images in a self-supervised way. The 3D shapes are represented continuously in function-space as signed distance functions which we pre-train from example shapes in a supervised way. These shape priors provide weak supervision signals to better condition the challenging overall learning task. We evaluate the accuracy of our model in inferring 3D scene layout, demonstrate its generative capabilities, assess its generalization to real images, and point out benefits of the learned representation.

📄 PDF Abstract BibTeX arXiv:2010.04030

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingScene UnderstandingWeakly-supervised Learning

Similar Papers 제목 키워드 기반

MONet: Unsupervised Scene Decomposition and Representation

2019-01-22 · Christopher P. Burgess, Loic Matthey, Nicholas Watters, Rishabh Kabra 외

The ability to decompose scenes in terms of abstract building blocks is crucial for general intelligence. Where those basic building blocks share meaningful properties, interactions and other regularities across scenes, …

Object DiscoveryUnsupervised Object Segmentation

Cycle Consistency in Video Object-Centric Learning

2026-05-28 · Rongzhen Zhao, Zhiyuan Li, Ruonan Wei, Juho Kannala 외 arxiv

Self-supervised video Object-Centric Learning (OCL) aims to discover distinct objects and associate them across time, whereas self-supervised Multi-Object Tracking (MOT) focuses on associating pre-defined object detectio…

Multi-Object Tracking

Weakly Supervised Image Annotation and Segmentation with Objects and Attributes

2017-08-08 · Zhiyuan Shi, Yongxin Yang, Timothy M. Hospedales, Tao Xiang

We propose to model complex visual scenes using a non-parametric Bayesian model learned from weakly labelled images abundant on media sharing sites such as Flickr. Given weak image-level annotations of objects and attrib…

AttributeObjectobject-detectionObject Detection+3

EgoFlowNet: Non-Rigid Scene Flow from Point Clouds with Ego-Motion Support

2024-07-03 · Ramy Battrawy, René Schuster, Didier Stricker

Recent weakly-supervised methods for scene flow estimation from LiDAR point clouds are limited to explicit reasoning on object-level. These methods perform multiple iterative optimizations for each rigid object, which ma…

ClusteringObjectScene Flow Estimation

Weakly Supervised 3D Object Detection from Lidar Point Cloud

2020-07-23 · ECCV 2020 8 · Qinghao Meng, Wenguan Wang, Tianfei Zhou, Jianbing Shen 외

It is laborious to manually label point cloud data for training high-quality 3D object detectors. This work proposes a weakly supervised approach for 3D object detection, only requiring a small set of weakly annotated sc…

3D Object DetectionObjectobject-detectionObject Detection