Papers Visual Relationship Detection
“Visual Relationship Detection” 태그가 달린 논문 82편 · 필터 해제
METOR: A Unified Framework for Mutual Enhancement of Objects and Relationships in Open-vocabulary Video Visual Relationship Detection
Open-vocabulary video visual relationship detection aims to detect objects and their relationships in videos without being restricted by predefined object or relationship categories. Existing methods leverage the rich se…
Objectobject-detectionObject DetectionRelationship Detection+1End-to-end Open-vocabulary Video Visual Relationship Detection using Multi-modal Prompting
Open-vocabulary video visual relationship detection aims to expand video visual relationship detection beyond annotated categories by detecting unseen relationships between both seen and unseen objects in videos. Existin…
DecoderObjectobject-detectionObject Detection+5Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
Visual Relationship Detection (VRD) has seen significant advancements with Transformer-based architectures recently. However, we identify two key limitations in a conventional label assignment for training Transformer-ba…
RelationRelationship DetectionScene Graph GenerationVisual Relationship DetectionScene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
Visual relationship detection aims to identify objects and their relationships in images. Prior methods approach this task by adding separate relationship modules or decoders to existing object detection architectures. T…
DecoderObjectobject-detectionObject Detection+2Video Relationship Detection Using Mixture of Experts
Machine comprehension of visual information from images and videos by neural networks faces two primary challenges. Firstly, there exists a computational and inference gap in connecting vision and language, making it dif…
Action RecognitionMixture-of-ExpertsObjectReading Comprehension+3RelVAE: Generative Pretraining for few-shot Visual Relationship Detection
Visual relations are complex, multimodal concepts that play an important role in the way humans perceive the world. As a result of their complexity, high-quality, diverse and large scale datasets for visual relations are…
Predicate ClassificationRelationship DetectionVisual Relationship DetectionSelf-Supervised Learning for Visual Relationship Detection through Masked Bounding Box Reconstruction
We present a novel self-supervised approach for representation learning, particularly for the task of Visual Relationship Detection (VRD). Motivated by the effectiveness of Masked Image Modeling (MIM), we propose Masked …
Predicate DetectionRelationship DetectionRepresentation LearningSelf-Supervised Learning+1STUPD: A Synthetic Dataset for Spatial and Temporal Relation Reasoning
Understanding relations between objects is crucial for understanding the semantics of a visual scene. It is also an essential step in order to bridge visual and language models. However, current state-of-the-art computer…
RelationRelationship DetectionSpatial ReasoningVisual Relationship DetectionNeSy4VRD: A Multifaceted Resource for Neurosymbolic AI Research using Knowledge Graphs in Visual Relationship Detection
NeSy4VRD is a multifaceted resource designed to support the development of neurosymbolic AI (NeSy) research. NeSy4VRD re-establishes public access to the images of the VRD dataset and couples them with an extensively rev…
Knowledge GraphsRelationship DetectionVisual Relationship DetectionUnified Visual Relationship Detection with Vision and Language Models
This work focuses on training a single visual relationship detector predicting over the union of label spaces from multiple datasets. Merging labels spanning different datasets could be challenging due to inconsistent ta…
Human-Object Interaction DetectionRelationship DetectionScene Graph GenerationVisual Relationship DetectionImage Semantic Relation Generation
Scene graphs provide structured semantic understanding beyond images. For downstream tasks, such as image retrieval, visual question answering, visual relationship detection, and even autonomous vehicle technology, scene…
Image RetrievalImage SegmentationImage to textQuestion Answering+8Distance-Aware Occlusion Detection with Focused Attention
For humans, understanding the relationships between objects using visual signals is intuitive. For artificial intelligence, however, this task remains challenging. Researchers have made significant progress studying sema…
DecoderHuman-Object Interaction DetectionRelationship DetectionVisual Relationship DetectionNeural Message Passing for Visual Relationship Detection
Visual relationship detection aims to detect the interactions between objects in an image; however, this task suffers from combinatorial explosion due to the variety of objects and interactions. Since the interactions as…
Relationship DetectionVisual Relationship DetectionLearning Structured Representations of Visual Scenes
As the intermediate-level representations bridging the two levels, structured representations of visual scenes, such as visual relationships between pairwise objects, have been shown to not only benefit compositional mod…
Human-Object Interaction DetectionRepresentation LearningScene Graph GenerationUnbiased Scene Graph Generation+1VReBERT: A Simple and Flexible Transformer for Visual Relationship Detection
Visual Relationship Detection (VRD) impels a computer vision model to 'see' beyond an individual object instance and 'understand' how different objects in a scene are related. The traditional way of VRD is first to detec…
ObjectRelationship DetectionVisual Relationship DetectionPEVL: Position-enhanced Pre-training and Prompt Tuning for Vision-language Models
Vision-language pre-training (VLP) has shown impressive performance on a wide range of cross-modal tasks, where VLP models without reliance on object detectors are becoming the mainstream due to their superior computatio…
Language ModelingLanguage ModellingObjectPhrase Grounding+6Scene Graph Generation: A Comprehensive Survey
Deep learning techniques have led to remarkable breakthroughs in the field of generic object detection and have spawned a lot of scene-understanding tasks in recent years. Scene graph has been the focus of research becau…
Graph Generationobject-detectionObject DetectionRelationship Detection+4A Probabilistic Graphical Model Based on Neural-Symbolic Reasoning for Visual Relationship Detection
This paper aims to leverage symbolic knowledge to improve the performance and interpretability of the Visual Relationship Detection (VRD) models. Existing VRD methods based on deep learning suffer from the problems o…
Deep LearningRelationship DetectionVisual Relationship DetectionRepresenting Prior Knowledge Using Randomly, Weighted Feature Networks for Visual Relationship Detection
The single-hidden-layer Randomly Weighted Feature Network (RWFN) introduced by Hong and Pavlic (2021) was developed as an alternative to neural tensor network approaches for relational learning tasks. Its relatively smal…
Predicate DetectionRelational ReasoningRelationship DetectionTensor Networks+2BGT-Net: Bidirectional GRU Transformer Network for Scene Graph Generation
Scene graphs are nodes and edges consisting of objects and object-object relationships, respectively. Scene graph generation (SGG) aims to identify the objects and their relationships. We propose a bidirectional GRU (BiG…
Graph GenerationObjectRelation PredictionRelationship Detection+2