paper-with-me

Papers

Shelf-Supervised Mesh Prediction in the Wild

2021-02-11 · CVPR 2021 1 · Yufei Ye, Shubham Tulsiani, Abhinav Gupta

We aim to infer 3D shape and pose of object from a single image and propose a learning-based approach that can train from unstructured image collections, supervised by only segmentation outputs from off-the-shelf recognition systems (i.e. 'shelf-supervised'). We first infer a volumetric representation in a canonical frame, along with the camera pose. We enforce the representation geometrically consistent with both appearance and masks, and also that the synthesized novel views are indistinguishable from image collections. The coarse volumetric prediction is then converted to a mesh-based representation, which is further refined in the predicted camera frame. These two steps allow both shape-pose factorization from image collections and per-instance reconstruction in finer details. We examine the method on both synthetic and real-world datasets and demonstrate its scalability on 50 categories in the wild, an order of magnitude more classes than existing works.

📄 PDF Abstract BibTeX arXiv:2102.06195

Code (1)

JudyYe/shelf-sup-mesh pytorch

Tasks

Prediction

Similar Papers 제목 키워드 기반

Learning 3D Human Dynamics from Video

2018-12-04 · CVPR 2019 6 · Angjoo Kanazawa, Jason Y. Zhang, Panna Felsen, Jitendra Malik

From an image of a person in action, we can easily guess the 3D motion of the person in the immediate past and future. This is because we have a mental model of 3D human dynamics that we have acquired from observing visu…

3D Human Dynamics3D Human Pose EstimationHuman Dynamics

Weakly-Supervised Mesh-Convolutional Hand Reconstruction in the Wild

2020-04-04 · CVPR 2020 6 · Dominik Kulon, Riza Alp Güler, Iasonas Kokkinos, Michael Bronstein 외

We introduce a simple and effective network architecture for monocular 3D hand pose estimation consisting of an image encoder followed by a mesh convolutional decoder that is trained through a direct 3D hand mesh reconst…

3D Hand Pose EstimationDecoderHand Pose EstimationPose Estimation

ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding

2025-12-03 · Lingjun Zhao, Yandong Luo, James Hays, Lu Gan arxiv

We introduce ShelfGaussian, an open-vocabulary multi-modal Gaussian-based 3D scene understanding framework supervised by off-the-shelf vision foundation models (VFMs). Gaussian-based methods have demonstrated superior pe…

Computational EfficiencyScene Understanding

MagicPony: Learning Articulated 3D Animals in the Wild

2022-11-22 · CVPR 2023 1 · Shangzhe Wu, Ruining Li, Tomas Jakab, Christian Rupprecht 외

We consider the problem of predicting the 3D shape, articulation, viewpoint, texture, and lighting of an articulated animal like a horse given a single test image as input. We present a new method, dubbed MagicPony, that…

Viewpoint Estimation

Online Adaptation for Consistent Mesh Reconstruction in the Wild

2020-12-06 · NeurIPS 2020 12 · Xueting Li, Sifei Liu, Shalini De Mello, Kihwan Kim 외

This paper presents an algorithm to reconstruct temporally consistent 3D meshes of deformable object instances from videos in the wild. Without requiring annotations of 3D mesh, 2D keypoints, or camera pose for each vide…

3D Reconstruction