paper-with-me

홈 › Papers

Image-to-Voxel Model Translation for 3D Scene Reconstruction and Segmentation

2020-08-01 · ECCV 2020 8 · Vladimir V. Kniaz, Vladimir A. Knyaz, Fabio Remondino, Artem Bordodymov, Petr Moshkantsev

Objects class, depth, and shape are instantly reconstructed by a human looking at a 2D image. While modern deep models solve each of these challenging tasks separately, they struggle to perform simultaneous scene 3D reconstruction and segmentation. We propose a single shot image-to-semantic voxel model translation framework. We train a generator adversarially against a discriminator that verifies the object's poses. Furthermore, trapezium-shaped voxels, volumetric residual blocks, and 2D-to-3D skip connections facilitate our model learning explicit reasoning about 3D scene structure. We collected a SemanticVoxels dataset with 116k images, ground-truth semantic voxel models, depth maps, and 6D object poses. Experiments on ShapeNet and our SemanticVoxels datasets demonstrate that our framework achieves and surpasses state-of-the-art in the reconstruction of scenes with multiple non-rigid objects of different classes. We made our model and dataset publicly available

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction3D Scene ReconstructionTranslation

Similar Papers 제목 키워드 기반

BUOL: A Bottom-Up Framework with Occupancy-aware Lifting for Panoptic 3D Scene Reconstruction From A Single Image

2023-06-01 · CVPR 2023 1 · Tao Chu, Pan Zhang, Qiong Liu, Jiaqi Wang

Understanding and modeling the 3D scene from a single image is a practical problem. A recent advance proposes a panoptic 3D scene reconstruction task that performs both 3D reconstruction and 3D panoptic segmentation from…

3D Panoptic Segmentation3D Reconstruction3D Scene ReconstructionPanoptic Segmentation

A One Stop 3D Target Reconstruction and multilevel Segmentation Method

2023-08-14 · Jiexiong Xu, Weikun Zhao, Zhiyan Tang, Xiangchao Gan

3D object reconstruction and multilevel segmentation are fundamental to computer vision research. Existing algorithms usually perform 3D scene reconstruction and target objects segmentation independently, and the perform…

3D Object Reconstruction3D Reconstruction3D Scene ReconstructionImage Segmentation+6

EPRecon: An Efficient Framework for Real-Time Panoptic 3D Reconstruction from Monocular Video

2024-09-03 · Zhen Zhou, Yunkai Ma, Junfeng Fan, Shaolin Zhang 외

Panoptic 3D reconstruction from a monocular video is a fundamental perceptual task in robotic scene understanding. However, existing efforts suffer from inefficiency in terms of inference speed and accuracy, limiting the…

3D ReconstructionScene Understanding

GS-Voxel: Fitting-Free Structured Latents for Large-Scale 3DGS Generation

2026-08-18 · Ming Qian, Zijian Wang, Minchao Sun, Jincheng Xiong 외 arxiv

Many scalable latent 3D generators operate on structured tensors, whereas pre-optimized 3D Gaussian Splatting (3DGS) reconstructions are unordered, spatially irregular, and vary widely in primitive count. We present GS-V…

Scene Generation

PanoSSC: Exploring Monocular Panoptic 3D Scene Reconstruction for Autonomous Driving

2024-06-11 · Yining Shi, Jiusi Li, Kun Jiang, Ke Wang 외

Vision-centric occupancy networks, which represent the surrounding environment with uniform voxels with semantics, have become a new trend for safe driving of camera-only autonomous driving perception systems, as they ar…

3D Instance Segmentation3D Scene Reconstruction3D Semantic SegmentationAutonomous Driving+5