paper-with-me

홈 › Papers

Weakly-supervised Disentangling with Recurrent Transformations for 3D View Synthesis

2016-01-05 · NeurIPS 2015 12 · Jimei Yang, Scott Reed, Ming-Hsuan Yang, Honglak Lee

An important problem for both graphics and vision is to synthesize novel views of a 3D object from a single image. This is particularly challenging due to the partial observability inherent in projecting a 3D object onto the image space, and the ill-posedness of inferring object shape and pose. However, we can train a neural network to address the problem if we restrict our attention to specific object categories (in our case faces and chairs) for which we can gather ample training data. In this paper, we propose a novel recurrent convolutional encoder-decoder network that is trained end-to-end on the task of rendering rotated objects starting from a single image. The recurrent structure allows our model to capture long-term dependencies along a sequence of transformations. We demonstrate the quality of its predictions for human faces on the Multi-PIE dataset and for a dataset of 3D chair models, and also show its ability to disentangle latent factors of variation (e.g., identity and pose) without using full supervision.

📄 PDF Abstract BibTeX arXiv:1601.00706

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderObject

Similar Papers 제목 키워드 기반

Quantum RNNs and LSTMs Through Entangling and Disentangling Power of Unitary Transformations

2025-05-10 · Ammar Daskin

In this paper, we discuss how quantum recurrent neural networks (RNNs) and their enhanced version, long short-term memory (LSTM) networks, can be modeled using the core ideas presented in Ref.[1], where the entangling an…

Recurrent Transformer Networks for Semantic Correspondence

2018-10-29 · NeurIPS 2018 12 · Seungryong Kim, Stephen Lin, Sangryul Jeon, Dongbo Min 외

We present recurrent transformer networks (RTNs) for obtaining dense correspondences between semantically similar images. Our networks accomplish this through an iterative process of estimating spatial transformations be…

General ClassificationSemantic correspondence

Dual Swap Disentangling

2018-05-27 · NeurIPS 2018 12 · Zunlei Feng, Xinchao Wang, Chenglong Ke, An-Xiang Zeng 외

Learning interpretable disentangled representations is a crucial yet challenging task. In this paper, we propose a weakly semi-supervised method, termed as Dual Swap Disentangling (DSD), for disentangling using both labe…

Attribute

Weakly-supervised 3D Pose Transfer with Keypoints

2023-07-25 · ICCV 2023 1 · Jinnan Chen, Chen Li, Gim Hee Lee

The main challenges of 3D pose transfer are: 1) Lack of paired training data with different characters performing the same pose; 2) Disentangling pose and shape information from the target mesh; 3) Difficulty in applying…

Pose Transfer

Discriminate-and-Rectify Encoders: Learning from Image Transformation Sets

2017-03-14 · Andrea Tacchetti, Stephen Voinea, Georgios Evangelopoulos

The complexity of a learning task is increased by transformations in the input space that preserve class identity. Visual object recognition for example is affected by changes in viewpoint, scale, illumination or planar …

Face VerificationObject RecognitionRetrievalWeakly-supervised Learning