paper-with-me

Papers Visual Relationship Detection

“Visual Relationship Detection” 태그가 달린 논문 82편 · 필터 해제

METOR: A Unified Framework for Mutual Enhancement of Objects and Relationships in Open-vocabulary Video Visual Relationship Detection

2025-05-10 · Yongqi Wang, Xinxiao wu, Shuo Yang

Open-vocabulary video visual relationship detection aims to detect objects and their relationships in videos without being restricted by predefined object or relationship categories. Existing methods leverage the rich se…

Objectobject-detectionObject DetectionRelationship Detection+1

End-to-end Open-vocabulary Video Visual Relationship Detection using Multi-modal Prompting

2024-09-19 · Yongqi Wang, Shuo Yang, Xinxiao wu, Jiebo Luo

Open-vocabulary video visual relationship detection aims to expand video visual relationship detection beyond annotated categories by detecting unseen relationships between both seen and unseen objects in videos. Existin…

DecoderObjectobject-detectionObject Detection+5

Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection

2024-03-26 · CVPR 2024 1 · Jongha Kim, Jihwan Park, Jinyoung Park, Jinyoung Kim 외

Visual Relationship Detection (VRD) has seen significant advancements with Transformer-based architectures recently. However, we identify two key limitations in a conventional label assignment for training Transformer-ba…

RelationRelationship DetectionScene Graph GenerationVisual Relationship Detection

Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection

2024-03-21 · Tim Salzmann, Markus Ryll, Alex Bewley, Matthias Minderer

Visual relationship detection aims to identify objects and their relationships in images. Prior methods approach this task by adding separate relationship modules or decoders to existing object detection architectures. T…

DecoderObjectobject-detectionObject Detection+2

Video Relationship Detection Using Mixture of Experts

2024-03-06 · IEEE Access 2023 3 · Ala Shaabana, Zahra Gharaee, Paul Fieguth

Machine comprehension of visual information from images and videos by neural networks faces two primary challenges. Firstly, there exists a computational and inference gap in connecting vision and language, making it dif…

Action RecognitionMixture-of-ExpertsObjectReading Comprehension+3

RelVAE: Generative Pretraining for few-shot Visual Relationship Detection

2023-11-27 · Sotiris Karapiperis, Markos Diomataris, Vassilis Pitsikalis

Visual relations are complex, multimodal concepts that play an important role in the way humans perceive the world. As a result of their complexity, high-quality, diverse and large scale datasets for visual relations are…

Predicate ClassificationRelationship DetectionVisual Relationship Detection

Self-Supervised Learning for Visual Relationship Detection through Masked Bounding Box Reconstruction

2023-11-08 · Zacharias Anastasakis, Dimitrios Mallis, Markos Diomataris, George Alexandridis 외

We present a novel self-supervised approach for representation learning, particularly for the task of Visual Relationship Detection (VRD). Motivated by the effectiveness of Masked Image Modeling (MIM), we propose Masked …

Predicate DetectionRelationship DetectionRepresentation LearningSelf-Supervised Learning+1

STUPD: A Synthetic Dataset for Spatial and Temporal Relation Reasoning

2023-09-13 · Palaash Agrawal, Haidi Azaman, Cheston Tan

Understanding relations between objects is crucial for understanding the semantics of a visual scene. It is also an essential step in order to bridge visual and language models. However, current state-of-the-art computer…

RelationRelationship DetectionSpatial ReasoningVisual Relationship Detection

NeSy4VRD: A Multifaceted Resource for Neurosymbolic AI Research using Knowledge Graphs in Visual Relationship Detection

2023-05-22 · David Herron, Ernesto Jiménez-Ruiz, Giacomo Tarroni, Tillman Weyde

NeSy4VRD is a multifaceted resource designed to support the development of neurosymbolic AI (NeSy) research. NeSy4VRD re-establishes public access to the images of the VRD dataset and couples them with an extensively rev…

Knowledge GraphsRelationship DetectionVisual Relationship Detection

Unified Visual Relationship Detection with Vision and Language Models

2023-03-16 · ICCV 2023 1 · Long Zhao, Liangzhe Yuan, Boqing Gong, Yin Cui 외

This work focuses on training a single visual relationship detector predicting over the union of label spaces from multiple datasets. Merging labels spanning different datasets could be challenging due to inconsistent ta…

Human-Object Interaction DetectionRelationship DetectionScene Graph GenerationVisual Relationship Detection

Image Semantic Relation Generation

2022-10-19 · Mingzhe Du

Scene graphs provide structured semantic understanding beyond images. For downstream tasks, such as image retrieval, visual question answering, visual relationship detection, and even autonomous vehicle technology, scene…

Image RetrievalImage SegmentationImage to textQuestion Answering+8

Distance-Aware Occlusion Detection with Focused Attention

2022-08-23 · Yang Li, Yucheng Tu, Xiaoxue Chen, Hao Zhao 외

For humans, understanding the relationships between objects using visual signals is intuitive. For artificial intelligence, however, this task remains challenging. Researchers have made significant progress studying sema…

DecoderHuman-Object Interaction DetectionRelationship DetectionVisual Relationship Detection

Neural Message Passing for Visual Relationship Detection

2022-08-08 · Yue Hu, Siheng Chen, Xu Chen, Ya zhang 외

Visual relationship detection aims to detect the interactions between objects in an image; however, this task suffers from combinatorial explosion due to the variety of objects and interactions. Since the interactions as…

Relationship DetectionVisual Relationship Detection

Learning Structured Representations of Visual Scenes

2022-07-09 · Meng-Jiun Chiou

As the intermediate-level representations bridging the two levels, structured representations of visual scenes, such as visual relationships between pairwise objects, have been shown to not only benefit compositional mod…

Human-Object Interaction DetectionRepresentation LearningScene Graph GenerationUnbiased Scene Graph Generation+1

VReBERT: A Simple and Flexible Transformer for Visual Relationship Detection

2022-06-18 · Yu Cui, Moshiur Farazi

Visual Relationship Detection (VRD) impels a computer vision model to 'see' beyond an individual object instance and 'understand' how different objects in a scene are related. The traditional way of VRD is first to detec…

ObjectRelationship DetectionVisual Relationship Detection

PEVL: Position-enhanced Pre-training and Prompt Tuning for Vision-language Models

2022-05-23 · Yuan YAO, Qianyu Chen, Ao Zhang, Wei Ji 외

Vision-language pre-training (VLP) has shown impressive performance on a wide range of cross-modal tasks, where VLP models without reliance on object detectors are becoming the mainstream due to their superior computatio…

Language ModelingLanguage ModellingObjectPhrase Grounding+6

Scene Graph Generation: A Comprehensive Survey

2022-01-03 · Guangming Zhu, Liang Zhang, Youliang Jiang, Yixuan Dang 외

Deep learning techniques have led to remarkable breakthroughs in the field of generic object detection and have spawned a lot of scene-understanding tasks in recent years. Scene graph has been the focus of research becau…

Graph Generationobject-detectionObject DetectionRelationship Detection+4

A Probabilistic Graphical Model Based on Neural-Symbolic Reasoning for Visual Relationship Detection

2022-01-01 · CVPR 2022 1 · Dongran Yu, Bo Yang, Qianhao Wei, Anchen Li 외

This paper aims to leverage symbolic knowledge to improve the performance and interpretability of the Visual Relationship Detection (VRD) models. Existing VRD methods based on deep learning suffer from the problems o…

Deep LearningRelationship DetectionVisual Relationship Detection

Representing Prior Knowledge Using Randomly, Weighted Feature Networks for Visual Relationship Detection

2021-11-20 · AAAI Workshop CLeaR 2022 2 · Jinyung Hong, Theodore P. Pavlic

The single-hidden-layer Randomly Weighted Feature Network (RWFN) introduced by Hong and Pavlic (2021) was developed as an alternative to neural tensor network approaches for relational learning tasks. Its relatively smal…

Predicate DetectionRelational ReasoningRelationship DetectionTensor Networks+2

BGT-Net: Bidirectional GRU Transformer Network for Scene Graph Generation

2021-09-11 · Naina Dhingra, Florian Ritter, Andreas Kunz

Scene graphs are nodes and edges consisting of objects and object-object relationships, respectively. Scene graph generation (SGG) aims to identify the objects and their relationships. We propose a bidirectional GRU (BiG…

Graph GenerationObjectRelation PredictionRelationship Detection+2
1–20 / 82 다음 →