FusionNet: 3D Object Classification Using Multiple Data Representations
High-quality 3D object recognition is an important component of many vision and robotics systems. We tackle the object recognition problem using two data representations, to achieve leading results on the Princeton ModelNet challenge. The two representations: 1. Volumetric representation: the 3D object is discretized spatially as binary voxels - $1$ if the voxel is occupied and $0$ otherwise. 2. Pixel representation: the 3D object is represented as a set of projected 2D pixel images. Current leading submissions to the ModelNet Challenge use Convolutional Neural Networks (CNNs) on pixel representations. However, we diverge from this trend and additionally, use Volumetric CNNs to bridge the gap between the efficiency of the above two representations. We combine both representations and exploit them to learn new features, which yield a significantly better classifier than using either of the representations in isolation. To do this, we introduce new Volumetric CNN (V-CNN) architectures.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Object Classification3D Object RecognitionClassificationGeneral ClassificationObjectObject RecognitionSimilar Papers 제목 키워드 기반
YOLO-FEDER FusionNet: A Novel Deep Learning Architecture for Drone Detection
Predominant methods for image-based drone detection frequently rely on employing generic object detection algorithms like YOLOv5. While proficient in identifying drones against homogeneous backgrounds, these algorithms o…
Objectobject-detectionObject DetectionBrainFusionNet: a deep learning and XAI model to understand local, global, and sequential features of MRI images for improved brain tumour detection
The noise of Magnetic Resonance Imaging MRI poses challenges for Deep Learning DL when tumor boundaries are obscured tumor location and appearance are complex Therefore we develop BrainFusionNet that combines Convolution…
Brain Tumor ClassificationImage ClassificationTransfer LearningGAF-FusionNet: Multimodal ECG Analysis via Gramian Angular Fields and Split Attention
Electrocardiogram (ECG) analysis plays a crucial role in diagnosing cardiovascular diseases, but accurate interpretation of these complex signals remains challenging. This paper introduces a novel multimodal framework(GA…
ECG ClassificationTime SeriesTime Series AnalysisPerformance Optimization of YOLO-FEDER FusionNet for Robust Drone Detection in Visually Complex Environments
Drone detection in visually complex environments remains challenging due to background clutter, small object scale, and camouflage effects. While generic object detectors like YOLO exhibit strong performance in low-textu…
Object DetectionFrustumFusionNets: A Three-Dimensional Object Detection Network Based on Tractor Road Scene
To address the issues of the existing frustum-based methods' underutilization of image information in road three-dimensional object detection as well as the lack of research on agricultural scenes, we constructed an obje…
Objectobject-detectionObject Detection