paper-with-me

Generalized Referring Expression Comprehension

1개 벤치마크 · 논문 10편 · 이 태스크의 논문 보기 →

Benchmarks

gRefCOCO

결과 5개

Most implemented

Papers

Making Dialogue Grounding Data Rich: A Three-Tier Data Synthesis Framework for Generalized Referring Expression Comprehension

2025-12-02 · Juexi Shao, Siyou Li, Yujian Gan, Chris Madge 외 arxiv

Dialogue-Based Generalized Referring Expression Comprehension (GREC) requires models to ground the expression and unlimited targets in complex visual scenes while resolving coreference across a long dialogue context. How…

Generalized Referring Expression Comprehension

LIHE: Linguistic Instance-Split Hyperbolic-Euclidean Framework for Generalized Weakly-Supervised Referring Expression Comprehension

2025-11-15 · Xianglong Shi, Silin Cheng, Sirui Zhao, Yunhan Jiang 외 arxiv

Existing Weakly-Supervised Referring Expression Comprehension (WREC) methods, while effective, are fundamentally limited by a one-to-one mapping assumption, hindering their ability to handle expressions corresponding to …

Generalized Referring Expression Comprehension

Improving Generalized Visual Grounding with Instance-aware Joint Learning

2025-09-17 · Ming Dai, Wenxuan Cheng, Jiang-Jiang Liu, Lingfeng Yang 외 arxiv

Generalized visual grounding tasks, including Generalized Referring Expression Comprehension (GREC) and Segmentation (GRES), extend the classical visual grounding paradigm by accommodating multi-target and non-target sce…

Generalized Referring Expression ComprehensionSemantic SegmentationVisual Grounding

Hierarchical Alignment-enhanced Adaptive Grounding Network for Generalized Referring Expression Comprehension

2025-01-02 · Yaxian Wang, Henghui Ding, Shuting He, Xudong Jiang 외

In this work, we address the challenging task of Generalized Referring Expression Comprehension (GREC). Compared to the classic Referring Expression Comprehension (REC) that focuses on single-target expressions, GREC ext…

Generalized Referring Expression ComprehensionGeneralized Referring Expression SegmentationObject CountingPhrase Grounding+3

SimVG: A Simple Framework for Visual Grounding with Decoupled Multi-modal Fusion

2024-09-26 · Ming Dai, Lingfeng Yang, Yihao Xu, ZhenHua Feng 외

Visual grounding is a common vision task that involves grounding descriptive sentences to the corresponding regions of an image. Most existing methods use independent image-text encoding and apply complex hand-crafted mo…

DescriptiveGeneralized Referring Expression ComprehensionReferring Expression ComprehensionVisual Grounding

GREC: Generalized Referring Expression Comprehension

2023-08-30 · Shuting He, Henghui Ding, Chang Liu, Xudong Jiang

The objective of Classic Referring Expression Comprehension (REC) is to produce a bounding box corresponding to the object mentioned in a given textual description. Commonly, existing datasets and techniques in classic R…

Generalized Referring Expression ComprehensionReferring ExpressionReferring Expression Comprehension

전체 10편 보기 →