paper-with-me

홈 › Papers

ShapeClipper: Scalable 3D Shape Learning from Single-View Images via Geometric and CLIP-based Consistency

2023-04-13 · CVPR 2023 1 · Zixuan Huang, Varun Jampani, Anh Thai, Yuanzhen Li, Stefan Stojanov, James M. Rehg

We present ShapeClipper, a novel method that reconstructs 3D object shapes from real-world single-view RGB images. Instead of relying on laborious 3D, multi-view or camera pose annotation, ShapeClipper learns shape reconstruction from a set of single-view segmented images. The key idea is to facilitate shape learning via CLIP-based shape consistency, where we encourage objects with similar CLIP encodings to share similar shapes. We also leverage off-the-shelf normals as an additional geometric constraint so the model can learn better bottom-up reasoning of detailed surface geometry. These two novel consistency constraints, when used to regularize our model, improve its ability to learn both global shape structure and local geometric details. We evaluate our method over three challenging real-world datasets, Pix3D, Pascal3D+, and OpenImages, where we achieve superior performance over state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2304.06247

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

Shape-Aware Human Pose and Shape Reconstruction Using Multi-View Images

2019-08-26 · ICCV 2019 10 · Junbang Liang, Ming C. Lin

We propose a scalable neural network framework to reconstruct the 3D mesh of a human body from multi-view images, in the subspace of the SMPL model. Use of multi-view images can significantly reduce the projection ambigu…

3D Human Pose EstimationMulti-view 3D Human Pose Estimation

Learning Pose-invariant 3D Object Reconstruction from Single-view Images

2020-04-03 · Bo Peng, Wei Wang, Jing Dong, Tieniu Tan

Learning to reconstruct 3D shapes using 2D images is an active research topic, with benefits of not requiring expensive 3D data. However, most work in this direction requires multi-view images for each object instance as…

3D Object ReconstructionDomain AdaptationObject Reconstruction

Single-view 3D Scene Reconstruction with High-fidelity Shape and Texture

2023-11-01 · Yixin Chen, Junfeng Ni, Nan Jiang, Yaowei Zhang 외

Reconstructing detailed 3D scenes from single-view images remains a challenging task due to limitations in existing approaches, which primarily focus on geometric shape recovery, overlooking object appearances and fine s…

3D Object Reconstruction3D Reconstruction3D scene Editing3D Scene Reconstruction+4

Photo-Geometric Autoencoding to Learn 3D Objects from Unlabelled Images

2019-06-04 · Shangzhe Wu, Christian Rupprecht, Andrea Vedaldi

We show that generative models can be used to capture visual geometry constraints statistically. We use this fact to infer the 3D shape of object categories from raw single-view images. Differently from prior work, we us…

Planes vs. Chairs: Category-guided 3D shape learning without any 3D cues

2022-04-21 · Zixuan Huang, Stefan Stojanov, Anh Thai, Varun Jampani 외

We present a novel 3D shape reconstruction method which learns to predict an implicit 3D shape representation from a single RGB image. Our approach uses a set of single-view images of multiple object categories without v…

3D Shape Reconstruction3D Shape RepresentationMetric Learning