paper-with-me

홈 › Papers

3D-A-Nets: 3D Deep Dense Descriptor for Volumetric Shapes with Adversarial Networks

2017-11-28 · Mengwei Ren, Liang Niu, Yi Fang

Recently researchers have been shifting their focus towards learned 3D shape descriptors from hand-craft ones to better address challenging issues of the deformation and structural variation inherently present in 3D objects. 3D geometric data are often transformed to 3D Voxel grids with regular format in order to be better fed to a deep neural net architecture. However, the computational intractability of direct application of 3D convolutional nets to 3D volumetric data severely limits the efficiency (i.e. slow processing) and effectiveness (i.e. unsatisfied accuracy) in processing 3D geometric data. In this paper, powered with a novel design of adversarial networks (3D-A-Nets), we have developed a novel 3D deep dense shape descriptor (3D-DDSD) to address the challenging issues of efficient and effective 3D volumetric data processing. We developed new definition of 2D multilayer dense representation (MDR) of 3D volumetric data to extract concise but geometrically informative shape description and a novel design of adversarial networks that jointly train a set of convolution neural network (CNN), recurrent neural network (RNN) and an adversarial discriminator. More specifically, the generator network produces 3D shape features that encourages the clustering of samples from the same category with correct class label, whereas the discriminator network discourages the clustering by assigning them misleading adversarial class labels. By addressing the challenges posed by the computational inefficiency of direct application of CNN to 3D volumetric data, 3D-A-Nets can learn high-quality 3D-DSDD which demonstrates superior performance on 3D shape classification and retrieval over other state-of-the-art techniques by a great margin.

📄 PDF Abstract BibTeX arXiv:1711.10108

Code (0)

등록된 구현이 없습니다.

Tasks

3D Shape ClassificationClusteringRetrieval

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

Learning a Probabilistic Latent Space of Object Shapes via 3D Generative-Adversarial Modeling

2016-10-24 · NeurIPS 2016 12 · Jiajun Wu, Chengkai Zhang, Tianfan Xue, William T. Freeman 외

We study the problem of 3D object generation. We propose a novel framework, namely 3D Generative Adversarial Network (3D-GAN), which generates 3D objects from a probabilistic space by leveraging recent advances in volume…

3D Object Recognition3D Point Cloud Linear ClassificationGenerative Adversarial NetworkObject+2

Dense Object Nets: Learning Dense Visual Object Descriptors By and For Robotic Manipulation

2018-06-22 · Peter R. Florence, Lucas Manuelli, Russ Tedrake

What is the right object representation for manipulation? We would like robots to visually perceive scenes and learn an understanding of the objects in them that (i) is task-agnostic and can be used as a building block f…

Object

Investigation of Densely Connected Convolutional Networks with Domain Adversarial Learning for Noise Robust Speech Recognition

2021-12-19 · Chia Yu Li, Ngoc Thang Vu

We investigate densely connected convolutional networks (DenseNets) and their extension with domain adversarial training for noise robust speech recognition. DenseNets are very deep, compact convolutional neural networks…

Robust Speech Recognitionspeech-recognitionSpeech Recognition

3D Topology Transformation with Generative Adversarial Networks

2020-07-07 · Luca Stornaiuolo, Nima Dehmamy, Albert-László Barabási, Mauro Martino

Generation and transformation of images and videos using artificial intelligence have flourished over the past few years. Yet, there are only a few works aiming to produce creative 3D shapes, such as sculptures. Here we …

Learning SO(3)-Invariant Semantic Correspondence via Local Shape Transform

2024-04-17 · CVPR 2024 1 · Chunghyun Park, SeungWook Kim, Jaesik Park, Minsu Cho

Establishing accurate 3D correspondences between shapes stands as a pivotal challenge with profound implications for computer vision and robotics. However, existing self-supervised methods for this problem assume perfect…

DecoderSemantic correspondence