Dual Octree Graph Networks for Learning Adaptive Volumetric Shape Representations
We present an adaptive deep representation of volumetric fields of 3D shapes and an efficient approach to learn this deep representation for high-quality 3D shape reconstruction and auto-encoding. Our method encodes the volumetric field of a 3D shape with an adaptive feature volume organized by an octree and applies a compact multilayer perceptron network for mapping the features to the field value at each 3D position. An encoder-decoder network is designed to learn the adaptive feature volume based on the graph convolutions over the dual graph of octree nodes. The core of our network is a new graph convolution operator defined over a regular grid of features fused from irregular neighboring octree nodes at different levels, which not only reduces the computational and memory cost of the convolutions over irregular neighboring octree nodes, but also improves the performance of feature learning. Our method effectively encodes shape details, enables fast 3D shape reconstruction, and exhibits good generality for modeling 3D shapes out of training categories. We evaluate our method on a set of reconstruction tasks of 3D shapes and scenes and validate its superiority over other existing approaches. Our code, data, and trained models are available at https://wang-ps.github.io/dualocnn.
Code (1)
Tasks
3D Shape ReconstructionDecoderMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Adaptive O-CNN: A Patch-based Deep Representation of 3D Shapes
We present an Adaptive Octree-based Convolutional Neural Network (Adaptive O-CNN) for efficient 3D shape encoding and decoding. Different from volumetric-based or octree-based CNN methods that represent a 3D shape with v…
DecoderOctree Generating Networks: Efficient Convolutional Architectures for High-resolution 3D Outputs
We present a deep convolutional decoder architecture that can generate volumetric 3D outputs in a compute- and memory-efficient manner by using an octree representation. The network learns to predict both the structure o…
3D ReconstructionDecoderVocal Bursts Intensity PredictionEfficient Autoregressive Shape Generation via Octree-Based Adaptive Tokenization
Many 3D generative models rely on variational autoencoders (VAEs) to learn compact shape representations. However, existing methods encode all shapes into a fixed-size token, disregarding the inherent variations in scale…
Open-Vocabulary Octree-Graph for 3D Scene Understanding
Open-vocabulary 3D scene understanding is indispensable for embodied agents. Recent works leverage pretrained vision-language models (VLMs) for object segmentation and project them to point clouds to build 3D maps. Despi…
ObjectScene UnderstandingSemantic SegmentationO-CNN: Octree-based Convolutional Neural Networks for 3D Shape Analysis
We present O-CNN, an Octree-based Convolutional Neural Network (CNN) for 3D shape analysis. Built upon the octree representation of 3D shapes, our method takes the average normal vectors of a 3D model sampled in the fine…
3D Object ClassificationGPURetrievalSemantic Segmentation