Papers Zero-Shot Human-Object Interaction Detection
“Zero-Shot Human-Object Interaction Detection” 태그가 달린 논문 9편 · 필터 해제
Locality-Aware Zero-Shot Human-Object Interaction Detection
Recent methods for zero-shot Human-Object Interaction (HOI) detection typically leverage the generalization ability of large Vision-Language Model (VLM), i.e., CLIP, on unseen categories, showing impressive results on va…
Human-Object Interaction DetectionObjectZero-Shot Human-Object Interaction DetectionUnseen No More: Unlocking the Potential of CLIP for Generative Zero-shot HOI Detection
Zero-shot human-object interaction (HOI) detector is capable of generalizing to HOI categories even not encountered during training. Inspired by the impressive zero-shot capabilities offered by CLIP, latest methods striv…
Human-Object Interaction DetectionZero-Shot Human-Object Interaction DetectionBoosting Zero-Shot Human-Object Interaction Detection with Vision-Language Transfer
Human-Object Interaction (HOI) detection is a crucial task that involves localizing interactive human-object pairs and identifying the actions being performed. Most existing HOI detectors are supervised in nature and lac…
Human-Object Interaction DetectionLanguage ModelingLanguage ModellingObject+1Towards Zero-shot Human-Object Interaction Detection via Vision-Language Integration
Human-object interaction (HOI) detection aims to locate human-object pairs and identify their interaction categories in images. Most existing methods primarily focus on supervised learning, which relies on extensive manu…
DecoderHuman-Object Interaction DetectionLanguage ModelingLanguage Modelling+2RLIPv2: Fast Scaling of Relational Language-Image Pre-training
Relational Language-Image Pre-training (RLIP) aims to align vision representations with relational texts, thereby advancing the capability of relational reasoning in computer vision tasks. However, hindered by the slow c…
Graph GenerationHuman-Object Interaction Detectionobject-detectionObject Detection+4Boosting Human-Object Interaction Detection with Text-to-Image Diffusion Model
This paper investigates the problem of the current HOI detection methods and introduces DiffHOI, a novel HOI detection scheme grounded on a pre-trained text-image diffusion model, which enhances the detector's performanc…
DiversityHuman-Object Interaction DetectionTripletZero-Shot Human-Object Interaction DetectionRelViT: Concept-guided Vision Transformer for Visual Relational Reasoning
Reasoning about visual relationships is central to how humans interpret the visual world. This task remains challenging for current deep learning algorithms since it requires addressing three key technical problems joint…
Human-Object Interaction DetectionObjectRetrievalSystematic Generalization+3End-to-End Zero-Shot HOI Detection via Vision and Language Knowledge Distillation
Most existing Human-Object Interaction~(HOI) Detection methods rely heavily on full annotations with predefined HOI categories, which is limited in diversity and costly to scale further. We aim at advancing zero-shot HOI…
Human-Object Interaction DetectionKnowledge DistillationObjectobject-detection+2ConsNet: Learning Consistency Graph for Zero-Shot Human-Object Interaction Detection
We consider the problem of Human-Object Interaction (HOI) Detection, which aims to locate and recognize HOI instances in the form of <human, action, object> in images. Most existing works treat HOIs as individual interac…
Human-Object Interaction DetectionObjectZero-Shot Human-Object Interaction Detection