Described Object Detection
1개 벤치마크 · 논문 9편 · 이 태스크의 논문 보기 →
Benchmarks
Most implemented
Grounded Language-Image Pre-training
Simple Open-Vocabulary Object Detection with Vision Transformers
SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models
Papers
What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging
State-of-the-art vision-language models (VLMs) suffer from a critical failure in understanding negation, often referred to as affirmative bias. This limitation is particularly severe in described object detection (DOD) t…
Described Object DetectionAn Open and Comprehensive Pipeline for Unified Object Grounding and Detection
Grounding-DINO is a state-of-the-art open-set detection model that tackles multiple vision tasks including Open-Vocabulary Detection (OVD), Phrase Grounding (PG), and Referring Expression Comprehension (REC). Its effecti…
Described Object DetectionPhrase GroundingReferring ExpressionReferring Expression ComprehensionSPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models
We present SPHINX, a versatile multi-modal large language model (MLLM) with a joint mixing of model weights, tuning tasks, and visual embeddings. First, for stronger vision-language alignment, we unfreeze the large langu…
Described Object DetectionLanguage ModelingLanguage ModellingLarge Language Model+4Described Object Detection: Liberating Object Detection with Flexible Expressions
Detecting objects based on language information is a popular task that includes Open-Vocabulary object Detection (OVD) and Referring Expression Comprehension (REC). In this paper, we advance them to a more practical sett…
Binary ClassificationDescribed Object DetectionObjectobject-detection+5CORA: Adapting CLIP for Open-Vocabulary Detection with Region Prompting and Anchor Pre-Matching
Open-vocabulary detection (OVD) is an object detection task aiming at detecting objects from novel categories beyond the base categories on which the detector is trained. Recent OVD methods rely on large-scale visual-lan…
Described Object Detectionobject-detectionObject DetectionObject Localization+1Universal Instance Perception as Object Discovery and Retrieval
All instance perception tasks aim at finding certain objects specified by some queries such as category names, language expressions, and target annotations, but this complete field has been split into multiple independen…
Described Object DetectionGeneralized Referring Expression ComprehensionInstance SegmentationMulti-Object Tracking and Segmentation+16