paper-with-me

홈 › Papers

ShapeCodes: Self-Supervised Feature Learning by Lifting Views to Viewgrids

2017-09-01 · ECCV 2018 9 · Dinesh Jayaraman, Ruohan Gao, Kristen Grauman

We introduce an unsupervised feature learning approach that embeds 3D shape information into a single-view image representation. The main idea is a self-supervised training objective that, given only a single 2D image, requires all unseen views of the object to be predictable from learned features. We implement this idea as an encoder-decoder convolutional neural network. The network maps an input image of an unknown category and unknown viewpoint to a latent space, from which a deconvolutional decoder can best "lift" the image to its complete viewgrid showing the object from all viewing angles. Our class-agnostic training procedure encourages the representation to capture fundamental shape primitives and semantic regularities in a data-driven manner---without manual semantic labels. Our results on two widely-used shape datasets show 1) our approach successfully learns to perform "mental rotation" even for objects unseen during training, and 2) the learned latent space is a powerful representation for object recognition, outperforming several existing unsupervised feature learning methods.

📄 PDF Abstract BibTeX arXiv:1709.00505

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderObjectObject Recognition

Similar Papers 제목 키워드 기반

Attention-Propagation Network for Egocentric Heatmap to 3D Pose Lifting

2024-02-28 · CVPR 2024 1 · Taeho Kang, Youngki Lee

We present EgoTAP, a heatmap-to-3D pose lifting method for highly accurate stereo egocentric 3D pose estimation. Severe self-occlusion and out-of-view limbs in egocentric camera views make accurate pose estimation a chal…

3D Pose EstimationEgocentric Pose EstimationPose EstimationPosition

C3G: Learning Compact 3D Representations with 2K Gaussians

2025-12-03 · Honggyu An, Jaewoo Jung, Mungyeom Kim, Chaehyun Kim 외 arxiv

Reconstructing and understanding 3D scenes from unposed sparse views in a feed-forward manner remains as a challenging task in 3D computer vision. Recent approaches use per-pixel 3D Gaussian Splatting for reconstruction,…

Novel View SynthesisScene Understanding

Unsupervised 3D Pose Estimation with Geometric Self-Supervision

2019-04-09 · CVPR 2019 6 · Ching-Hang Chen, Ambrish Tyagi, Amit Agrawal, Dylan Drover 외

We present an unsupervised learning approach to recover 3D human pose from 2D skeletal joints extracted from a single image. Our method does not require any multi-view image data, 3D skeletons, correspondences between 2D…

3D Human Pose Estimation3D Pose EstimationPose Estimationvalid

Feature-Suppressed Contrast for Self-Supervised Food Pre-training

2023-08-07 · Xinda Liu, Yaohui Zhu, Linhu Liu, Jiang Tian 외

Most previous approaches for analyzing food images have relied on extensively annotated datasets, resulting in significant human labeling expenses due to the varied and intricate nature of such images. Inspired by the ef…

Food RecognitionSelf-Supervised Learning

Self-supervised Discriminative Feature Learning for Deep Multi-view Clustering

2021-03-28 · Jie Xu, Yazhou Ren, Huayi Tang, Zhimeng Yang 외

Multi-view clustering is an important research topic due to its capability to utilize complementary information from multiple views. However, there are few methods to consider the negative impact caused by certain views …

ClusteringDiversity