paper-with-me

Papers

Improving Visual Relation Detection using Depth Maps

2019-05-02 · Sahand Sharifzadeh, Sina Moayed Baharlou, Max Berrendorf, Rajat Koner, Volker Tresp

Visual relation detection methods rely on object information extracted from RGB images such as 2D bounding boxes, feature maps, and predicted class probabilities. We argue that depth maps can additionally provide valuable information on object relations, e.g. helping to detect not only spatial relations, such as standing behind, but also non-spatial relations, such as holding. In this work, we study the effect of using different object features with a focus on depth maps. To enable this study, we release a new synthetic dataset of depth maps, VG-Depth, as an extension to Visual Genome (VG). We also note that given the highly imbalanced distribution of relations in VG, typical evaluation metrics for visual relation detection cannot reveal improvements of under-represented relations. To address this problem, we propose using an additional metric, calling it Macro Recall@K, and demonstrate its remarkable performance on VG. Finally, our experiments confirm that by effective utilization of depth maps within a simple, yet competitive framework, the performance of visual relation detection can be improved by a margin of up to 8%.

📄 PDF Abstract BibTeX arXiv:1905.00966

Code (1)

Sina-Baharlou/Depth-VRD 공식 구현 pytorch

Tasks

ObjectRelationRelationship DetectionVisual Relationship Detection

Similar Papers 제목 키워드 기반

A Study on the Relationship Between Depth Map Quality and the Overall 3D Video Quality OF Experience

2018-03-14

The emergence of multiview displays has made the need for synthesizing virtual views more pronounced, since it is not practical to capture all of the possible views when filming multiview content. View synthesis is perfo…

Correlation of Object Detection Performance with Visual Saliency and Depth Estimation

2024-11-05 · Matthias Bartolo, Dylan Seychell

As object detection techniques continue to evolve, understanding their relationships with complementary visual tasks becomes crucial for optimising model architectures and computational resources. This paper investigates…

Depth EstimationDepth PredictionFeature EngineeringObject+4

VCP-DCN: Beyond Visual Concealed Property via Depth Collaborative Network for Camouflaged Object Detection

2026-07-30 · Songsong Duan, Xi Yang, Nannan Wang arxiv

Camouflaged Object Detection (COD) aims to identify and segment camouflaged objects in complex environments, which are often concealed because their color and texture are similar to the background. Several existing COD m…

Contrastive LearningObject Detection

DistillGrasp: Integrating Features Correlation with Knowledge Distillation for Depth Completion of Transparent Objects

2024-08-01 · Yiheng Huang, Junhong Chen, Nick Michiels, Muhammad Asim 외

Due to the visual properties of reflection and refraction, RGB-D cameras cannot accurately capture the depth of transparent objects, leading to incomplete depth maps. To fill in the missing points, recent studies tend to…

Depth CompletionFeature CorrelationKnowledge DistillationRobotic Grasping+1

DepthGAN: GAN-based Depth Generation of Indoor Scenes from Semantic Layouts

2022-03-22 · Yidi Li, Yiqun Wang, Zhengda Lu, Jun Xiao

Limited by the computational efficiency and accuracy, generating complex 3D scenes remains a challenging problem for existing generation networks. In this work, we propose DepthGAN, a novel method of generating depth map…

Computational Efficiency