paper-with-me

Visual Relationship Detection

5개 벤치마크 · 논문 82편 · 이 태스크의 논문 보기 →

Benchmarks

VRD Phrase Detection

결과 28개

VRD

결과 4개

Visual Genome

결과 4개

Most implemented

Papers

METOR: A Unified Framework for Mutual Enhancement of Objects and Relationships in Open-vocabulary Video Visual Relationship Detection

2025-05-10 · Yongqi Wang, Xinxiao wu, Shuo Yang

Open-vocabulary video visual relationship detection aims to detect objects and their relationships in videos without being restricted by predefined object or relationship categories. Existing methods leverage the rich se…

Objectobject-detectionObject DetectionRelationship Detection+1

End-to-end Open-vocabulary Video Visual Relationship Detection using Multi-modal Prompting

2024-09-19 · Yongqi Wang, Shuo Yang, Xinxiao wu, Jiebo Luo

Open-vocabulary video visual relationship detection aims to expand video visual relationship detection beyond annotated categories by detecting unseen relationships between both seen and unseen objects in videos. Existin…

DecoderObjectobject-detectionObject Detection+5

Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection

2024-03-26 · CVPR 2024 1 · Jongha Kim, Jihwan Park, Jinyoung Park, Jinyoung Kim 외

Visual Relationship Detection (VRD) has seen significant advancements with Transformer-based architectures recently. However, we identify two key limitations in a conventional label assignment for training Transformer-ba…

RelationRelationship DetectionScene Graph GenerationVisual Relationship Detection

Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection

2024-03-21 · Tim Salzmann, Markus Ryll, Alex Bewley, Matthias Minderer

Visual relationship detection aims to identify objects and their relationships in images. Prior methods approach this task by adding separate relationship modules or decoders to existing object detection architectures. T…

DecoderObjectobject-detectionObject Detection+2

Video Relationship Detection Using Mixture of Experts

2024-03-06 · IEEE Access 2023 3 · Ala Shaabana, Zahra Gharaee, Paul Fieguth

Machine comprehension of visual information from images and videos by neural networks faces two primary challenges. Firstly, there exists a computational and inference gap in connecting vision and language, making it dif…

Action RecognitionMixture-of-ExpertsObjectReading Comprehension+3

RelVAE: Generative Pretraining for few-shot Visual Relationship Detection

2023-11-27 · Sotiris Karapiperis, Markos Diomataris, Vassilis Pitsikalis

Visual relations are complex, multimodal concepts that play an important role in the way humans perceive the world. As a result of their complexity, high-quality, diverse and large scale datasets for visual relations are…

Predicate ClassificationRelationship DetectionVisual Relationship Detection

전체 82편 보기 →