paper-with-me

Described Object Detection

1개 벤치마크 · 논문 9편 · 이 태스크의 논문 보기 →

Benchmarks

Most implemented

Grounded Language-Image Pre-training

2021-12-07 · 구현 3개

Papers

What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging

2025-10-15 · Inha Kang, Youngsun Lim, Seonho Lee, Jiho Choi 외 arxiv

State-of-the-art vision-language models (VLMs) suffer from a critical failure in understanding negation, often referred to as affirmative bias. This limitation is particularly severe in described object detection (DOD) t…

Described Object Detection

An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

2024-01-04 · Xiangyu Zhao, Yicheng Chen, Shilin Xu, Xiangtai Li 외

Grounding-DINO is a state-of-the-art open-set detection model that tackles multiple vision tasks including Open-Vocabulary Detection (OVD), Phrase Grounding (PG), and Referring Expression Comprehension (REC). Its effecti…

Described Object DetectionPhrase GroundingReferring ExpressionReferring Expression Comprehension

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

2023-11-13 · Ziyi Lin, Chris Liu, Renrui Zhang, Peng Gao 외

We present SPHINX, a versatile multi-modal large language model (MLLM) with a joint mixing of model weights, tuning tasks, and visual embeddings. First, for stronger vision-language alignment, we unfreeze the large langu…

Described Object DetectionLanguage ModelingLanguage ModellingLarge Language Model+4

Described Object Detection: Liberating Object Detection with Flexible Expressions

2023-07-24 · NeurIPS 2023 11 · Chi Xie, Zhao Zhang, Yixuan Wu, Feng Zhu 외

Detecting objects based on language information is a popular task that includes Open-Vocabulary object Detection (OVD) and Referring Expression Comprehension (REC). In this paper, we advance them to a more practical sett…

Binary ClassificationDescribed Object DetectionObjectobject-detection+5

CORA: Adapting CLIP for Open-Vocabulary Detection with Region Prompting and Anchor Pre-Matching

2023-03-23 · CVPR 2023 1 · Xiaoshi Wu, Feng Zhu, Rui Zhao, Hongsheng Li

Open-vocabulary detection (OVD) is an object detection task aiming at detecting objects from novel categories beyond the base categories on which the detector is trained. Recent OVD methods rely on large-scale visual-lan…

Described Object Detectionobject-detectionObject DetectionObject Localization+1

Universal Instance Perception as Object Discovery and Retrieval

2023-03-12 · CVPR 2023 1 · Bin Yan, Yi Jiang, Jiannan Wu, Dong Wang 외

All instance perception tasks aim at finding certain objects specified by some queries such as category names, language expressions, and target annotations, but this complete field has been split into multiple independen…

Described Object DetectionGeneralized Referring Expression ComprehensionInstance SegmentationMulti-Object Tracking and Segmentation+16

전체 9편 보기 →