paper-with-me

홈 › Papers

Learning Common Representation from RGB and Depth Images

2018-12-17 · Giorgio Giannone, Boris Chidlovskii

We propose a new deep learning architecture for the tasks of semantic segmentation and depth prediction from RGB-D images. We revise the state of art based on the RGB and depth feature fusion, where both modalities are assumed to be available at train and test time. We propose a new architecture where the feature fusion is replaced with a common deep representation. Combined with an encoder-decoder type of the network, the architecture can jointly learn models for semantic segmentation and depth estimation based on their common representation. This representation, inspired by multi-view learning, offers several important advantages, such as using one modality available at test time to reconstruct the missing modality. In the RGB-D case, this enables the cross-modality scenarios, such as using depth data for semantically segmentation and the RGB images for depth estimation. We demonstrate the effectiveness of the proposed network on two publicly available RGB-D datasets. The experimental results show that the proposed method works well in both semantic segmentation and depth estimation tasks.

📄 PDF Abstract BibTeX arXiv:1812.06873

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDepth EstimationDepth PredictionMULTI-VIEW LEARNINGSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

RCDPT: Radar-Camera fusion Dense Prediction Transformer

2022-11-04 · Chen-Chou Lo, Patrick Vandewalle

Recently, transformer networks have outperformed traditional deep neural networks in natural language processing and show a large potential in many computer vision tasks compared to convolutional backbones. In the origin…

Depth EstimationMonocular Depth EstimationPrediction

Depth Is All You Need for Monocular 3D Detection

2022-10-05 · Dennis Park, Jie Li, Dian Chen, Vitor Guizilini 외

A key contributor to recent progress in 3D detection from single images is monocular depth estimation. Existing methods focus on how to leverage depth explicitly, by generating pseudo-pointclouds or providing attention c…

AllDepth EstimationDepth PredictionMonocular Depth Estimation+1

Geometry meets semantics for semi-supervised monocular depth estimation

2018-10-09 · Pierluigi Zama Ramirez, Matteo Poggi, Fabio Tosi, Stefano Mattoccia 외

Depth estimation from a single image represents a very exciting challenge in computer vision. While other image-based depth sensing techniques leverage on the geometry between different viewpoints (e.g., stereo or struct…

DecoderDepth EstimationDepth PredictionMonocular Depth Estimation+1

NeuralODF: Learning Omnidirectional Distance Fields for 3D Shape Representation

2022-06-12 · Trevor Houchens, Cheng-You Lu, Shivam Duggal, Rao Fu 외

In visual computing, 3D geometry is represented in many different forms including meshes, point clouds, voxel grids, level sets, and depth images. Each representation is suited for different tasks thus making the transfo…

3D geometry3D Shape Representation

Learning With Side Information Through Modality Hallucination

2016-06-01 · CVPR 2016 6 · Judy Hoffman, Saurabh Gupta, Trevor Darrell

We present a modality hallucination architecture for training an RGB object detection model which incorporates depth side information at training time. Our convolutional hallucination network learns a new and complementa…

Hallucinationobject-detectionObject Detection