Beam Search for Learning a Deep Convolutional Neural Network of 3D Shapes
This paper addresses 3D shape recognition. Recent work typically represents a 3D shape as a set of binary variables corresponding to 3D voxels of a uniform 3D grid centered on the shape, and resorts to deep convolutional neural networks(CNNs) for modeling these binary variables. Robust learning of such CNNs is currently limited by the small datasets of 3D shapes available, an order of magnitude smaller than other common datasets in computer vision. Related work typically deals with the small training datasets using a number of ad hoc, hand-tuning strategies. To address this issue, we formulate CNN learning as a beam search aimed at identifying an optimal CNN architecture, namely, the number of layers, nodes, and their connectivity in the network, as well as estimating parameters of such an optimal CNN. Each state of the beam search corresponds to a candidate CNN. Two types of actions are defined to add new convolutional filters or new convolutional layers to a parent CNN, and thus transition to children states. The utility function of each action is efficiently computed by transferring parameter values of the parent CNN to its children, thereby enabling an efficient beam search. Our experimental evaluation on the 3D ModelNet dataset demonstrates that our model pursuit using the beam search yields a CNN with superior performance on 3D shape classification than the state of the art.
Code (1)
Tasks
3D Shape Classification3D Shape RecognitionSimilar Papers 제목 키워드 기반
Identification of multiple damage in beams based on robust curvature mode shapes
Multiple damage identification in beams using curvature mode shape has become a research focus of increasing interest during the last few years. On this topic, most existing studies address the sensitivity of curvature…
SensitivityGPU-Accelerated Inverse Lithography Towards High Quality Curvy Mask Generation
Inverse Lithography Technology (ILT) has emerged as a promising solution for photo mask design and optimization. Relying on multi-beam mask writers, ILT enables the creation of free-form curvilinear mask shapes that enha…
GPUJointly optimal dereverberation and beamforming
We previously proposed an optimal (in the maximum likelihood sense) convolutional beamformer that can perform simultaneous denoising and dereverberation, and showed its superiority over the widely used cascade of a WPE d…
DenoisingUncertainty Determines the Adequacy of the Mode and the Tractability of Decoding in Sequence-to-Sequence Models
In many natural language processing (NLP) tasks the same input (e.g. source sentence) can have multiple possible outputs (e.g. translations). To analyze how this ambiguity (also known as intrinsic uncertainty) shapes the…
Grammatical Error CorrectionMachine TranslationSentenceNear-field Beam Steering with Planar Antenna Array
Beam steering enables manipulation of the electromagnetic radiation patterns in antenna array systems. A methodology for steering beams in the near field of a planar antenna array with known phase wavefront functions tow…